PubMed Health⌕ Search

SEARCH · PubMed Health

Results for “Form Perception”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 433 records · Page 24Linked to original sources

Additive effect of luminance and color cues in generation of neon color spreading.

A series of experiments were designed to examine how luminance and color cues influence the occurrence of neon color spreading for the Ehrenstein pattern plus the cross pattern. The proportion of "see" responses for color spreading was obtained for different combinations of the luminance and color of the pattern components. The following were obtained. (1) Increase in the luminance of the inducing pattern and/or the color difference between the cross and the inducing pattern raised the proportion of "see" responses for color spreading. This implies that luminance and color signals additively contribute to the generation of the color spreading, but in different ways. (2) The "iso-spreading contours" for the generation of color spreading were determined on the two-dimensional isoluminant plane composed of the L-M and S-(L+M) axes. The contours were approximately described by rotated quadratic or ellipse functions. The additive interaction within the chromatic systems [the L-M system and the S-(L+M) system] was less significant than that across the luminance and chromatic systems. (3) The pattern of the experimental results could not be explained straightforwardly by existing models.

Color Perception↗

Integrating contours within and through depth.

To better understand the role of disparity in contour integration we compared detection performance of "paths" composed of elements confined either to a single depth plane, or spanning multiple depth planes. In both cases paths defined by alignment of elements were embedded in a noise background-field made up of similar, but randomly positioned, elements covering the same depth range as the path elements. We show that a systematic disparity cue can enhance the detectability of paths which traverse depth, but that this detectability is weak compared to paths made up of elements of the same disparity. These results suggest that the outputs of disparity detectors tuned to different disparities can be linked to define contours.

Depth Perception↗

Light source dependence in shape from shading.

We investigated shape constancy in human shape from shading under variations of illuminant direction using a local attitude probe in conjunction with a perturbation analysis. Stimuli were computer generated and depicted ellipsoids in a structured setting. Even with these simple shapes subjects settings were systematically biased in the illuminant direction and were consistent with a regression to image luminance gradients. These biases were reduced for high albedo scenes where interreflections make image illuminance more dependent on scene geometry. Adding texture to the surface reduced but did not eliminate this bias. These results suggest that we can expect little constancy in shape from shading under variations of illuminant direction without constraints from other cues.

Computer Simulation↗

2D motion aliasing yielding 3D ambiguity. A study with variants of a Necker cube.

The 2D projection of a rotating Necker cube yields an ambiguous 3D interpretation based on both 2D shape and kinetic depth information. The present study shows that the alternation rate of the two 3D interpretations is constant with the rotation speed up to some critical value (around 25 turns/min for a cube whose sides subtend 2.5 deg) and increases monotonically thereafter. It is proposed that the additional perceptual reversals (PRs) observed at high rotation speeds are due to the increased frequency of the crossovers of the cube's edges. These crossovers yield 2D motion "aliasing" (or discontinuity) and "veridical" (or continuity) motion components. The motion aliasing (or crossover) hypothesis states that, in addition to the inherent ambiguity of the dynamic 2D projection of 3D objects, perceptual motion/perspective reversals will occur any time the discontinuity speed takes over the continuity speed. It is proposed that the relative strengths of the two components depend on the linear speed of the projected edges and that the discontinuity components take over the continuity one in the speed range where contrast sensitivity (or, above threshold, efficiency) is a decreasing function of speed. The motion aliasing hypothesis was tested and supported in a series of independent experiments showing that, for rotation speeds higher than 25 turns/min the PR rate increases with the crossover frequency at a constant speed, with linear speed at a constant crossover frequency and with the similarity of the crossing bars in terms of their orientation, polarity and spatial overlap. In addition, some of these experiments suggest that 2D shape and kinetic depth 3D-cues combine in such a way that the average PR rate they yield together is the same as the PR rate yielded by each of them independently. In the Discussion section we elaborate on issues related to the perceptual combination of ambiguous shape and kinetic depth, 3D cues.

Depth Perception↗

Concentric orientation summation in human form vision.

Psychophysical data demonstrate that orientation information in concentric, random-dot Glass patterns is summed linearly to extract a global form percept. Surprisingly, no such global pooling was found for Glass patterns with parallel structure. A simple neural model explains these results and agrees with recent V4 single unit physiology. As V4 provides the major input to IT, global concentric units may play an important role in analyzing complex images such as faces. In support of this possibility, deficits in the perception of concentric Glass patterns have recently been linked to prosopagnosia.

Humans↗

The viewpoint complexity of an object-recognition task.

There is an ongoing debate about the nature of perceptual representation in human object recognition. Resolution of this debate has been hampered by the lack of a metric for assessing the representational requirements of a recognition task. To recognize a member of a given set of 3-D objects, how much detail must the objects' representations contain in order to achieve a specific accuracy criterion? From the performance of an ideal observer, we derived a quantity called the view complexity (VX) to measure the required granularity of representation. VX is an intrinsic property of the object-recognition task, taking into account both the object ensemble and the type of decision required of an observer. It does not depend on the visual representation or processing used by the observer. VX can be interpreted as the number of randomly selected 2-D images needed to represent the decision boundaries in the image space of a 3-D object-recognition task. A low VX means the task is inherently more viewpoint invariant and a high VX means it is inherently more viewpoint dependent. By measuring the VX of recognition tasks with different object sets, we show that the current confusion about the nature of human perceptual representation is partly due to a failure in distinguishing between human visual processing and the properties of a task and its stimuli. We find general correspondence between the VX of a recognition task and the published human data on viewpoint dependence. Exceptions in this relationship motivated us to propose the view-rate hypothesis: human visual performance is limited by the equivalent number of 2-D image views that can be processed per unit time.

Algorithms↗

Local and global factors affecting the coherent motion of gratings presented in multiple apertures.

Using stimuli composed of two independent gratings viewed through multiple apertures, we investigate a number of parameters affecting the integration of locally ambiguous motions into globally coherent motion. In four experiments, we varied local factors (grating spatial frequency, speed, contrast, duty cycle, orientation) and global factors (degree of similarity and common fate between the gratings, and symmetry in the configuration of the grating pattern) and examined their effects on global motion coherence. Our results, confirming accounts offered by previous investigators, indicate that local competition between motion signals generated by contours (ambiguous) and their line terminations (unambiguous) is important in determining global motion coherence in multiple-aperture stimuli. Our results also indicate that global factors can affect perceived coherence independently of local motion signals, suggesting the involvement of higher-level motion areas and a role for non-motion processes such as those involved in pattern and form perception. Comparing motion coherence with other two-dimensional (2-D) stimuli (plaids) shows that 2-D multiple-aperture stimuli are not analogous and that coherence models derived from plaid stimuli do not account for the data.

Humans↗

Shape from shading: estimation of reflectance map.

The reflectance map used by the visual system for perception of shape from shading was estimated. In Experiment 1, an image of a cylinder or a sphere illuminated from the viewer's direction was presented, and subjects estimated the cross-section of perceived 3D-shape. The reflectance map was estimated from the relationship between the stimulus image intensities and the slants of the measured cross-section. The estimated reflectance maps were not the ones based on Lambertian reflectance properties. In Experiment 2, whether perceived shapes could be predicted based on the reflectance maps obtained in Experiment 1 was examined. Subjects performed the same shape estimation task with images of cylinders generated by the reflectance map obtained in Experiment 1. The perceived shapes coincided well with the shapes used for stimulus image generation. These results indicate that the visual system's estimation of shape from shading can be fully understood based on empirically obtained reflectance maps without mentioning its inaccurate nature which has been claimed by past studies.

Algorithms↗

Large-scale tests of a keyed, appearance-based 3-D object recognition system.

We describe and analyze an appearance-based 3-D object recognition system that avoids some of the problems of previous appearance-based schemes. We describe various large-scale performance tests and report good performance for full-sphere/hemisphere recognition of up to 24 complex, curved objects, robustness against clutter and occlusion, and some intriguing generic recognition behavior. We also establish a protocol that permits performance in the presence of quantifiable amounts of clutter and occlusion to be predicted on the basis of simple score statistics derived from clean test images and pure clutter images.

Computer Simulation↗

Quantitative depth for a phantom surface can be based on cyclopean occlusion cues alone.

Liu, L., Stevenson, S.B., and Schor, C.M. (1994, Nature, 367, 66-669) reported quantitative stereoscopic depth in a phantom rectangle which appeared to lack conventional matching elements. Later, Gillam, B.J. (1995, Nature, 373, 202-203) and Liu, L., Stevenson, S.B., and Schor, C.M. (1995, Nature, 373, 203) and Liu, L., Stevenson, S.B., and Schor, C.M. (1997, Vision Research, 37(5), 633-644) indicated that the varying depth of the phantom rectangle could be based on stereoscopic matching. To remove the contaminating effects of conventional stereopsis from the Liu et al. (1994) original example, we presented a pair of parallel vertical lines to each eye where there is a central gap in the right line for the left eye's view and in the left line for the right eye's view. Observers saw a phantom rectangle bounded by subjective contours whose depth increased with the thickness of the lines. We attribute the quantitative variation of depth to a purely cyclopean (binocular) process sensitive to the pattern of contour presence and absence in the two eye's view.

Cues↗

2D observers for human 3D object recognition?

In human object recognition, converging evidence has shown that subjects' performance depends on their familiarity with an object's appearance. The extent of such dependence is a function of the inter-object similarity. The more similar the objects are, the stronger this dependence will be and the more dominant the two-dimensional (2D) image-based information will be. However, the degree to which three-dimensional (3D) model-based information is used remains an area of strong debate. Previously the authors showed that all models with independent 2D templates that allowed 2D rotations in the image plane cannot account for human performance in discriminating novel object views. Here the authors derive an analytic formulation of a Bayesian model that gives rise to the best possible performance under 2D affine transformations and demonstrate that this model cannot account for human performance in 3D object discrimination. Relative to this model, human statistical efficiency is higher for novel views than for learned views, suggesting that human observers have used some 3D structural information.

Computer Simulation↗

Stereopsis based on monocular gaps: metrical encoding of depth and slant without matching contours.

It is often the case in binocular vision that one eye can see between two objects lying at different distances but the other eye cannot. We have found that the visual system is able to correctly interpret images produced this way in which a single solid rectangle in one eye is fused with two half-sized rectangles in the other eye separated by a vertical gap comprising the background. Two rectangles in depth are seen. It is as if the solid rectangle is treated as two components which each match one of the physically separated rectangles in the contralateral eye. The sign of the depth depends on which eye's view has the gap and its magnitude increases with gap width. Measured depth is found to be equivalent to real stereoscopic depth with a relative disparity equal to the monocular gap. If overall disparity differences are eliminated, between the left and the right images, variations in perceived slant of the two rectangles are still seen with increasing gap size. That two surfaces can be seen in metric binocular depth despite complete camouflage of their separation in one eye's view, suggests that stereopsis be regarded as a broad process of surface recovery not necessarily requiring image disparity at the location of the depth step.

Depth Perception↗

Categorical learning in pigeons: the role of texture and shape in complex static stimuli.

Pigeons are known to be able to categorize a wide variety of visual stimulus classes. However, it remains unclear which are the characteristics of the perceptually relevant features employed to reach such good performance. Here, we investigate the relative contributions of texture and shape information to categorization decisions about complex natural classes. We trained three groups of pigeons to discriminate between sets of photorealistic frontal images of human faces according to sex and subsequently, tested them on different stimulus sets. Only the pigeons that were presented with texture information were successful at the discrimination task. Pigeons seem to possess a sophisticated texture processing system but are less capable in discriminating shapes. The results are discussed in terms of the possible evolutionary advantages of utilizing texture as a very general and potent perceptual dimension in the birds' visual environment.

Animals↗

Perceived distance, shape and size.

If distance, shape and size are judged independently from the retinal and extra-retinal information at hand, different kinds of information can be expected to dominate each judgement, so that errors in one judgement need not be consistent with errors in other judgements. In order to evaluate how independent these three judgments are, we examined how adding information that improves one judgement influences the others. Subjects adjusted the size and the global shape of a computer-simulated ellipsoid to match a tennis ball. They then indicated manually where they judged the simulated ball to be. Adding information about distance improved the three judgements in a consistent manner, demonstrating that a considerable part of the errors in all three judgements were due to misestimating the distance. Adding information about shape that is independent of distance improved subjects' judgements of shape, but did not influence the set size or the manually indicated distance. Thus, subjects ignored conflicts between the cues when judging the shape, rather than using the conflicts to improve their estimate of the ellipsoid's distance. We conclude that the judgements are quite independent, in the sense that no attempt is made to attain consistency, but that they do rely on some common measures, such as that of distance.

Cues↗

Feature specific segmentation in perceived structure-from-motion.

Motion information is important to vision for extracting the 3-D (three-dimensional) structure of an object, as evidenced by the compelling percept of three-dimensionality attainable in displays which are purely motion-defined. It has recently been shown that when subjects view a rotating transparent cylinder of dots simulated with parallel projection, they rarely perceive rotation reversals which are physically introduced (Treue, Andersen, Ando & Hildreth, Vision Research, 35;1995:139-148). We show however that when the elements defining the cylinder are oriented, the number of perceived reversals increases systematically to near maximum as the difference between element orientations on the two surfaces increases. These results imply that structure-from-motion mechanisms are capable of exploiting local feature differences between the different surfaces of a moving object.

Depth Perception↗

One-shot viewpoint invariance in matching novel objects.

Humans often evidence little difficulty at recognizing objects from arbitrary orientations in depth. According to one class of theories, this competence is based on generalization from templates specified by metric properties (MPs), that were learned for the various orientations. An alternative class of theories assumes that non-accidental properties (NAPs) might be exploited so that even novel objects can be recognized under depth rotation. After scaling MP and NAP differences so that they were equally detectable when the objects were at the same orientation in depth, the present investigation assessed the effects of rotation on same-different judgments for matching novel objects. Judgments of a sequential pair of images of novel objects, when rendered from different viewpoints, revealed relatively low costs when the objects differed in a NAP of a single part, i.e. a geon. However, rotation dramatically reduced the detectability of MP differences to a level well below that expected by chance. NAPs offer a striking advantage over MPs for object classification and are therefore more likely to play a central role in the representation of objects.

Adolescent↗

Interaction between the perceived shape of two objects.

The difference between the way in which binocular disparity scales with viewing distance and the way in which motion parallax scales with viewing distance introduces a potential indirect cue for viewing distance: the viewing distance is the only distance at which disparity and motion specify the same depth. The present study examines whether this information is used. Two simulated ellipsoids were presented on a computer screen in complete darkness. The two ellipsoids were 6 degrees to the left and right of straight ahead. Subjects set the width and depth of each ellipsoid to match a tennis ball, and set the distance of the one on the right to half that of the one on the left. The distance of the left ellipsoid varied between trials. On half of the trials it was static. On the other half it was rotating up and down around its frontal horizontal axis. Rotating the left ellipsoid influenced its set depth: rotating ellipsoids were set to be much more spherical. There was no influence on the set depth of the other ellipsoid, or on the set width of either. The set distance of the right ellipsoid was also unaffected. We conclude that subjects do not combine binocular disparity and motion parallax to obtain more veridical information about viewing distance.

Distance Perception↗

Perception of three-dimensional shape from texture is based on patterns of oriented energy.

This paper presents empirical support for a new observer model of inferring three-dimensional shape from monocular texture cues. By measuring observers' abilities to estimate the relative three-dimensional curvature along a textured surface from two-dimensional projected images, and concurrently examining the local spectral changes occurring in the projected image for various texture patterns, we have found that correlated changes in oriented energy along lines corresponding to the lines of maximum and minimum curvature of the surface are crucial for conveying the three-dimensional shape of the surface. Energy along these lines of maximum and minimum curvature can be used to compute the orientation of local surface patches. Texture patterns consisting of simple and complex sinusoidal gratings and plaids, and filtered noise were drawn onto a surface that was corrugated sinusoidally in depth about the horizontal axis and projected in perspective onto an image plane. The perceived relative surface curvature was reconstructed from measurements of local ordinal depth around a central fixation point at 12 different phases of the corrugation. Our results show that: (1) it is neither necessary nor sufficient to identify individual texture elements or texture gradients in order to extract the shape of the surface; (2) one-dimensional frequency modulation is insufficient for conveying complex three-dimensional shape. (3) Veridical ordinal depth is seen only when the projected pattern contains changes in oriented energy along lines corresponding to projected lines of maximum curvature of the surface. (4) For a surface corrugated in depth about the horizontal axis, this pattern of oriented energy arises from energy along the vertical direction in the global Fourier transform of the pre-corrugated pattern. (5) Local orientation changes across lines of minimum curvature can be also critical for conveying shape. (6) These correlated orientation changes along lines of maximum and minimum curvature are entirely lost in parallel projection. Hence texture is a useful cue for shape if the image is a perspective projection. (7) Only some natural textures will provide sufficient monocular cues to support veridical shape inferences, and this can be predicted from their global Fourier transforms.

Computer Simulation↗