PubMed Health⌕ Search

PubMed · 11587722

Segmentation in structure from motion: modeling and psychophysics.

Abstract

Much work has been done on the question of how the visual system extracts the three-dimensional (3D) structure and motion of an object from two-dimensional (2D) motion information, a problem known as 'Structure from Motion', or SFM. Much less is known, however, about the human ability to recover structure and motion when the optic flow field arises from multiple objects, although observations of this ability date as early as Ullman's well-known two-cylinders stimulus [The interpretation of visual motion (1979)]. In the presence of multiple objects, the SFM problem is further aggravated by the need to solve the segmentation problem, i.e. deciding which motion signal belongs to which object. Here, we present a model for how the human visual system solves the combined SFM and segmentation problems, which we term SSFM, concurrently. The model is based on computation of a simple scalar property of the optic flow field known as def, which was previously shown to be used by human observers in SFM. The def values of many triplets of moving dots are computed, and the identification of multiple objects the image is based on detecting multiple peaks in the histogram of def values. In five experiments, we show that human SSFM performance is consistent with the predictions of the model. We compare the predictions of our model to those of other theoretical approaches, in particular those that use a rigidity hypothesis, and discuss the validity of each approach as a model for human SSFM.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

C Caudek, N Rubin. 2001. Segmentation in structure from motion: modeling and psychophysics.. https://doi.org/10.1016/s0042-6989(01)00163-8

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Discrimination training alters object representations in human extrastriate cortex.

Visual object recognition relies critically on learning. However, little is known about the effect of object learning in human visual cortex, and in particular how the spatial distribution of training effects relates to the distribution of object and face selectivity across the cortex before training. We scanned human subjects with high-resolution functional magnetic resonance imaging (fMRI) while they viewed novel object classes, both before and after extensive training to discriminate between exemplars within one of these object classes. Training increased the strength of the response in visual cortex to trained objects compared with untrained objects. However, training did not simply induce a uniform increase in the response to trained objects: the magnitude of this training effect varied substantially across subregions of extrastriate cortex, with some showing a twofold increase in response to trained objects and others (including the right fusiform face area) showing no significant effect of training. Furthermore, the spatial distribution of training effects could not be predicted from the spatial distribution of either pretrained responses or face selectivity. Instead, training changed the spatial distribution of activity across the cortex. These findings support a dynamic view of the ventral visual pathway in which the cortical representation of an object category is continuously modulated by experience.

Discrimination, Psychological↗

3D shape discrimination using relative disparity derivatives.

Three-dimensional (3D) shape discrimination could be achieved using relative disparity signals or it could be achieved using a higher-order disparity derivative detector. Two 3D shape discrimination tasks were used to distinguish between these possibilities: a within-shape task and a between-shape task. Disparity thresholds were larger when discriminating within the same shape than when discriminating between shapes. More importantly, within-shape discriminations were dependent on the pedestal disparity (distance from fixation) whereas between-shape discriminations were not. The results suggest that a mechanism sensitive to higher-order disparity derivatives can achieve discrimination between different 3D shapes.

Discrimination, Psychological↗

Learning alters local face space geometry.

The effects of learning on the geometry of face space were investigated by measuring thresholds for discrimination and recognition of synthetic faces. This was based on a novel experimental technique that permitted measurement of psychometric functions for face recognition. Two major results were obtained. First, thresholds for face recognition were significantly better than thresholds for discrimination among novel faces. Second, rapid discrimination in the neighborhood of learned faces was better than discrimination near novel faces. Control experiments showed that this discrimination improvement occurred only with learned faces, and it could not be explained by generalized discrimination learning. Thus, face learning selectively alters or distorts face space in the vicinity of learned faces. This alteration may be due to an improvement in the signal/noise ratio as a result of face learning.

Discrimination, Psychological↗