PubMed Health⌕ Search

PubMed · 11195306

Phase congruency: a low-level image invariant.

Abstract

Phase congruency is a low-level invariant property of image features. Interest in low-level image invariants has been limited. This is surprising, considering the fundamental importance of being able to obtain reliable results from low-level image operations in order to successfully perform any higher level operations. However, an impediment to the use of phase congruency to detect features has been its sensitivity to noise. This paper extends the theory behind the calculation of phase congruency in a number of ways. An effective method of noise compensation is presented that only assumes that the noise power spectrum is approximately constant. Problems with the localization of features are addressed by introducing a new, more sensitive measure of phase congruency. The existing theory that has been developed for 1D signals is extended to allow the calculation of phase congruency in 2D images. Finally, it is argued that high-pass filtering should be used to obtain image information at different scales. With this approach, the choice of scale only affects the relative significance of features without degrading their localization.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

P Kovesi. 2000. Phase congruency: a low-level image invariant.. https://doi.org/10.1007/s004260000024

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Absence of flash-lag when judging global shape from local positions.

When a flash is presented aligned with a moving stimulus, the former is perceived to lag behind the latter (the flash-lag effect). We study whether this mislocalization occurs when a positional judgment is not required, but a veridical spatial relationship between moving and flashed stimuli is needed to perceive a global shape. To do this, we used Glass patterns that are formed by pairs of correlated dots. One dot of each pair was presented moving and, at a given moment, the other dot of each pair was flashed in order to build the Glass pattern. If a flash-lag effect occurs between each pair of dots, we expect the best perception of the global shape to occur when the flashed dots are presented before the moving dots arrive at the position that physically builds the Glass pattern. Contrary to this, we found that the best detection of Glass patterns occurred for the situation of physical alignment. This result is not consistent with a low-level contribution to the flash-lag effect.

Form Perception↗

Object recognition and segmentation by a fragment-based hierarchy.

How do we learn to recognize visual categories, such as dogs and cats? Somehow, the brain uses limited variable examples to extract the essential characteristics of new visual categories. Here, I describe an approach to category learning and recognition that is based on recent computational advances. In this approach, objects are represented by a hierarchy of fragments that are extracted during learning from observed examples. The fragments are class-specific features and are selected to deliver a high amount of information for categorization. The same fragments hierarchy is then used for general categorization, individual object recognition and object-parts identification. Recognition is also combined with object segmentation, using stored fragments, to provide a top-down process that delineates object boundaries in complex cluttered scenes. The approach is computationally effective and provides a possible framework for categorization, recognition and segmentation in human vision.

Form Perception↗

Dynamics of shape interaction in human vision.

Spatial context can alter perceived shape, and temporal context can influence the perception of a stimulus. We sought to determine the time course of shape interactions by using a paradigm in which closed shape contours are laterally displaced over space and time. Target and masks are separated by various stimulus onset asynchrony (SOA) values, yielding forward, backward, and simultaneous masking conditions. Results indicate that spatial lateral interactions of shape are amplified by temporal asynchrony, reaching a peak at SOAs of 80-110 ms. Mask amplitude scales all effects and masking is shape specific. When a single mask follows the target, both spatial configuration and mask onset transient are critical in determining depth of masking. When the target is followed by two sequential masks, the possibility of apparent motion determines whether one or both masks drive masking. These findings suggest that temporal interactions of shape are dependent on an interactive combination of shape specificity and transients, that apparent motion plays a modulatory role, and that target shape is determined after a temporal window, not at its onset.

Form Perception↗