PubMed HealthSearch

PubMed · 9304683

Perceptual categories for spatial layout.

Abstract

The central problems of vision are often divided into object identification and localization. Object identification, at least at fine levels of discrimination, may require the application of top-down knowledge to resolve ambiguous image information. Utilizing top-down knowledge, however, may require the initial rapid access of abstract object categories based on low-level image cues. Does object localization require a different set of operating principles than object identification or is category determination also part of the perception of depth and spatial layout? Three-dimensional graphics movies of objects and their cast shadows are used to argue that identifying perceptual categories is important for determining the relative depths of objects. Processes that can identify the causal class (e.g. the kind of material) that generates the image data can provide information to determine the spatial relationships between surfaces. Changes in the blurriness of an edge may be characteristically associated with shadows caused by relative motion between two surfaces. The early identification of abstract events such as moving object/shadow pairs may also be important for depth from shadows. Knowledge of how correlated motion in the image relates to an object and its shadow may provide a reliable cue to access such event categories.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

D Kersten. 1997-08-29. Perceptual categories for spatial layout.. https://doi.org/10.1098/rstb.1997.0099

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Comparative study of two-dimensional and three-dimensional vision systems for minimally invasive surgery.

BACKGROUND: The aim of this comparative study was to gain subjective and objective data to determine for which operative tasks three-dimensional (3-D) vision systems are superior to two-dimensional (2-D) systems and to demonstrate any advantages or disadvantages of 3-D systems. METHODS: A model with five standardized tasks including sewing and knotting was developed to objectively measure performance times and to count technical faults. In our training center for minimally invasive surgery, surgeons involved in basic and advanced laparoscopic courses trained using both 2-D and 3-D vision systems. They subsequently completed analog scale questionnaires to record a subjective impression of comparative ease of operation tasks under 2-D and 3-D vision and to identify perceived deficiencies in the 3-D system. RESULTS: Compared to 2-D vision, the objective performance time was significantly shorter and significantly less mistakes were made using 3-D vision. All operative tasks were subjectively judged significantly easier under 3-D vision. CONCLUSIONS: Users with a normal capability for spatial perception can perform standard tasks more quickly and safely using 3-D vision, and a greater benefit is apparent for more complicated surgical maneuvers.

Depth Perception

Stereoscopic segregation of transparent surfaces and the effect of motion contrast.

Stereoscopic segregation in depth was studied using two superimposed frontoparallel surfaces displayed in dynamic random dot stereograms. The two patterns were positioned symmetrically in front of and behind a binocular fixation point. They were either stationary, or they could move relative to each other. Sensitivity for segregation was established by adding gaussian distributed disparity noise to the disparities specifying the two planes, and finding the noise amplitude that gave threshold segregation performance. Observers easily segregate the two surfaces for disparity differences between approximately 6 and 30-40 arcmin. Motion contrast, which by itself provides no cue to perform the task, greatly improves sensitivity for segregation. Noise tolerance rises by a factor of two or more when the patterns move at different speeds, or in different (frontoparallel) directions. The effect increases with directional difference, but the optimal directional difference deviated from 180 deg. The optimal speed varies with disparity difference. Thus, motion and disparity must interact in order to resolve the two transparent planes.

Depth Perception

dmax for stereopsis and motion in random dot displays.

The upper displacement limit for motion was compared with the upper disparity limit for stereopsis using two-frame random dot kinematograms or briefly presented stereograms. dmax (the disparity/displacement at which subjects make 20% errors in a forced-choice paradigm) was found to be very similar for motion and stereo at all dot densities, and to fall with increasing dot density (0.006% or two dots to 50%) according to a power law (exponent -0.2). If dmax is limited by the spacing of false targets, this pattern of results suggests that the spatial primitives in the input to the correspondence process may be derived from multiple spatial scales. A model using MIRAGE centroids provides a good fit to the data.

Depth Perception