PubMed HealthSearch

PubMed · 8423265

A model for context effects in speech recognition.

Abstract

A model is presented that quantifies the effect of context on speech recognition. In this model, a speech stimulus is considered as a concatenation of a number of equivalent elements (e.g., phonemes constituting a word). The model employs probabilities that individual elements are recognized and chances that missed elements are guessed using contextual information. Predictions are given of the probability that the entire stimulus, or part of it, is reproduced correctly. The model can be applied to both speech recognition and visual recognition of printed text. It has been verified with data obtained with syllables of the consonant-vowel-consonant (CVC) type presented near the reception threshold in quiet and in noise, with the results of an experiment using orthographic presentation of incomplete CVC syllables and with results of word counts in a CVC lexicon. A remarkable outcome of the analysis is that the cues which occur only in spoken language (e.g., coarticulatory cues) seem to have a much greater influence on recognition performance when the stimuli are presented near the threshold in noise than when they are presented near the absolute threshold. Demonstrations are given of further predictions provided by the model: word recognition as a function of signal-to-noise ratio, closed-set word recognition, recognition of interrupted speech, and sentence recognition.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

A W Bronkhorst, A J Bosman, G F Smoorenburg. 1993. A model for context effects in speech recognition.. https://doi.org/10.1121/1.406844

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Independent control of head and gaze movements during head-free pursuit in humans.

1. Head and gaze movements are usually highly co-ordinated. Here we demonstrate that under certain circumstances they can be controlled independently and we investigate the role of anticipatory activity in this process. 2. In experiment 1, subjects tracked, with head and eyes, a sinusoidally moving target. Overall, head and gaze trajectories were tightly coupled. From moment to moment, however, the trajectories could be very different and head movements were significantly more variable than gaze movements. 3. Predictive head and gaze responses can be elicited by repeated presentation of an intermittently illuminated, constant velocity target. In experiment 2 this protocol elicited a build-up of anticipatory head and gaze velocity, in opposing directions, when subjects made head movements in the opposite direction to target movement whilst maintaining gaze on target. 4. In experiment 3, head and gaze movements were completely uncoupled. Subjects followed, with head and gaze, respectively, two targets moving at different, harmonically unrelated frequencies. This was possible when both targets were visual, and also when gaze followed a visual target at one frequency whilst the head was oscillated in time with an auditory tone modulated at the second frequency. 5. We conclude that these results provide evidence of a visuomotor predictive mechanism that continuously samples visual feedback information and stores it such that it can be accessed by either the eye or the head to generate anticipatory movements. This overcomes time delays in visuomotor processing and facilitates time-sharing of motor activities, making possible the performance of two tasks simultaneously.

Acoustic Stimulation

Effects of stimulus frequency and intensity on c-fos mRNA expression in the adult rat auditory brainstem.

Induction of the cellular fos gene (c-fos) is one of the earliest transcriptional changes observed following neuronal excitation. Although not an activity marker in the strict electrophysiological sense, many neurons in the central nervous system increase their c-fos expression after periods of sustained stimulation at physiological levels of intensity. In the present study, induction of c-fos mRNA expression was examined in the auditory brainstem after 1 hour of continuous free-field acoustic stimulation. Sprague-Dawley rats were exposed to pure tones of 2, 8, 16, or 32 kHz or half-octave noise bands centered on 2, 8, or 32 kHz at 80-120 dB SPL. Stimulation-induced c-fos mRNA expression was evident at all levels of the auditory brainstem, and this expression was intensity dependent. In some brain areas, induced expression manifested a clear tonotopic organization, i.e., in dorsal, posteroventral, and anteroventral cochlear nuclei, and in the medial nucleus of the trapezoid body. The inferior colliculus exhibited multiple tonotopic representations. The dorsal nucleus of the lateral lemniscus had a crude tonotopy. Although expression was present, tonotopy was not evident in periolivary nuclei or in the ventral or intermediate nuclei of the lateral lemniscus. Free-field diotic stimulation did not induce c-fos mRNA expression in the medial or lateral superior olivary nuclei. Expression was induced in the lateral superior olive by dichotic stimulation (after a unilateral cochlear ablation), and that expression was tonotopically organized. The results suggest that stimulation-induced c-fos mRNA expression can be an effective way of mapping neuronal activity in the central auditory system under both normal and pathological conditions.

Acoustic Stimulation

Cochlear ablation alters acoustically induced c-fos mRNA expression in the adult rat auditory brainstem.

Expression of c-fos mRNA was studied in the adult rat brain following cochlear ablations by using in situ hybridization. In normal animals, expression was produced by acoustic stimulation and was found to be tonotopically distributed in many auditory nuclei. Following unilateral cochlear ablation, acoustically driven expression was eliminated or decreased in areas normally activated by the ablated ear, e.g., the ipsilateral dorsal and ventral cochlear nuclei, dorsal periolivary nuclei, and lateral nucleus of the trapezoid body and the contralateral medial and ventral nuclei of the trapezoid body, lateral lemniscal nuclei, and inferior colliculus. These deficits did not recover, even after long survivals up to 6 months. Results also indicated that neurons in the dorsal cochlear nucleus could be activated by contralateral stimulation in the absence of ipsilateral cochlear input and that the influence of the contralateral ear was tonotopically organized. Results also indicated that c-fos expression rose rapidly and persisted for up to 6 months in neurons in the rostral part of the contralateral medial nucleus of the trapezoid body following a cochlear ablation, even in the absence of acoustic stimulation. This response may reflect a release of constitutive excitatory inputs normally suppressed by missing afferent input or changes in homeostatic gene expression related to sensory deprivation. Instances of transient, surgery-dependent increases in c-fos mRNA expression in the absence of acoustic stimulation were observed in the superficial dorsal cochlear nucleus and the cochlear nerve root on the ablated side.

Acoustic Stimulation