PubMed Health⌕ Search

PubMed · 10694959

Temporal constraints on visual learning: a computational model.

Abstract

Given a constant stream of perceptual stimuli, how can the underlying invariances associated with a given input be learned? One approach consists of using generic truths about the spatiotemporal structure of the physical world as constraints on the types of quantities learned. The learning methodology employed here embodies one such truth: that perceptually salient properties (such as stereo disparity) tend to vary smoothly over time. Unfortunately, the units of an artificial neural network tend to encode superficial image properties, such as individual grey-level pixel values, which vary rapidly over time. However, if the states of units are constrained to vary slowly, then the network is forced to learn a smoothly varying function of the training data. We implemented this temporal-smoothness constraint in a backpropagation network which learned stereo disparity from random-dot stereograms. Temporal smoothness was formalized with the use of regularization theory by modifying the standard cost function minimised during training of a network. Temporal smoothness was found to be similar to other techniques for improving generalisation, such as early stopping and weight decay. However, in contrast to these, the theoretical underpinnings of temporal smoothing are intimately related to fundamental characteristics of the physical world. Results are discussed in terms of regularization theory and the physically realistic assumptions upon which temporal smoothing is based.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

J V Stone, N Harper. 1999. Temporal constraints on visual learning: a computational model.. https://doi.org/10.1068/p281089

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Simulated self-motion alters perceived time to collision.

Many authors have assumed that motor actions required for collision avoidance and for collision achievement (for example, in driving a car or hitting a ball) are guided by monitoring the time to collision (TTC), and that this is done on the basis of moment-to-moment values of the optical variable tau [1] [2] [3]. This assumption has also motivated the search for single neurons that fire when tau is a certain value [4] [5] [6] [7] [8]. Almost all of the laboratory studies and all the animal experiments were restricted to the case of stationary observer and moving object. On the face of it, this would seem reasonable. Even though humans and other animals routinely perform visually guided actions that require the TTC of an approaching object to be estimated while the observer is moving, tau provides an accurate estimate of TTC regardless of whether the approach is produced by self-motion, object-motion or a combination of both. One might therefore expect that judgements of TTC would be independent of self-motion. We report here, however, that simulated selfmotion using a peripheral flow field substantially altered estimates of TTC for an approaching object, even though the peripheral flow field did not affect the value of tau for the approaching object. This finding points to long range interactions between collision-sensitive visual neurons and neural mechanisms for processing self-motion.

Depth Perception↗