PubMed Health⌕ Search

PubMed · 15971688

Efficient visual search without top-down or bottom-up guidance.

Abstract

Two types of mechanisms have dominated theoretical accounts of efficient visual search. The first are bottom-up processes related to the characteristics of retinotopic feature maps. The second are top-down mechanisms related to feature selection. To expose the potential involvement of other mechanisms, we introduce a new search paradigm whereby a target is defined only in a context-dependent manner by multiple conjunctions of feature dimensions. Because targets in a multiconjunction task cannot be distinguished from distractors either by bottom-up guidance or top-down guidance, current theories of visual search predict inefficient search. While inefficient search does occur for the multiple conjunctions of orientation with color or luminance, we find efficient search for multiple conjunctions of luminance/size, luminance/shape, and luminance/topology. We also show that repeated presentations of either targets or a set of distractors result in much faster performance and that bottom-up feature extraction and top-down selection cannot account for efficient search on their own. In light of this, we discuss the possible role of perceptual organization in visual search. Furthermore, multiconjunction search could provide a new method for investigating perceptual grouping in visual search.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

DeLiang Wang, Arni Kristjansson, Ken Nakayama. 2005. Efficient visual search without top-down or bottom-up guidance.. https://doi.org/10.3758/bf03206488

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Absence of flash-lag when judging global shape from local positions.

When a flash is presented aligned with a moving stimulus, the former is perceived to lag behind the latter (the flash-lag effect). We study whether this mislocalization occurs when a positional judgment is not required, but a veridical spatial relationship between moving and flashed stimuli is needed to perceive a global shape. To do this, we used Glass patterns that are formed by pairs of correlated dots. One dot of each pair was presented moving and, at a given moment, the other dot of each pair was flashed in order to build the Glass pattern. If a flash-lag effect occurs between each pair of dots, we expect the best perception of the global shape to occur when the flashed dots are presented before the moving dots arrive at the position that physically builds the Glass pattern. Contrary to this, we found that the best detection of Glass patterns occurred for the situation of physical alignment. This result is not consistent with a low-level contribution to the flash-lag effect.

Form Perception↗

Object recognition and segmentation by a fragment-based hierarchy.

How do we learn to recognize visual categories, such as dogs and cats? Somehow, the brain uses limited variable examples to extract the essential characteristics of new visual categories. Here, I describe an approach to category learning and recognition that is based on recent computational advances. In this approach, objects are represented by a hierarchy of fragments that are extracted during learning from observed examples. The fragments are class-specific features and are selected to deliver a high amount of information for categorization. The same fragments hierarchy is then used for general categorization, individual object recognition and object-parts identification. Recognition is also combined with object segmentation, using stored fragments, to provide a top-down process that delineates object boundaries in complex cluttered scenes. The approach is computationally effective and provides a possible framework for categorization, recognition and segmentation in human vision.

Form Perception↗

Dynamics of shape interaction in human vision.

Spatial context can alter perceived shape, and temporal context can influence the perception of a stimulus. We sought to determine the time course of shape interactions by using a paradigm in which closed shape contours are laterally displaced over space and time. Target and masks are separated by various stimulus onset asynchrony (SOA) values, yielding forward, backward, and simultaneous masking conditions. Results indicate that spatial lateral interactions of shape are amplified by temporal asynchrony, reaching a peak at SOAs of 80-110 ms. Mask amplitude scales all effects and masking is shape specific. When a single mask follows the target, both spatial configuration and mask onset transient are critical in determining depth of masking. When the target is followed by two sequential masks, the possibility of apparent motion determines whether one or both masks drive masking. These findings suggest that temporal interactions of shape are dependent on an interactive combination of shape specificity and transients, that apparent motion plays a modulatory role, and that target shape is determined after a temporal window, not at its onset.

Form Perception↗