PubMed Health⌕ Search

SEARCH · PubMed Health

Results for “Speech Perception”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2Linked to original sources

An audio-visual corpus for speech perception and automatic speech recognition.

An audio-visual corpus has been collected to support the use of common material in speech perception and automatic speech recognition studies. The corpus consists of high-quality audio and video recordings of 1000 sentences spoken by each of 34 talkers. Sentences are simple, syntactically identical phrases such as "place green at B 4 now". Intelligibility tests using the audio signals suggest that the material is easily identifiable in quiet and low levels of stationary noise. The annotated corpus is available on the web for research use.

Acoustic Stimulation↗

The influence of phonemic awareness development on acoustic cue weighting strategies in children's speech perception.

In speech perception, children give particular patterns of weight to different acoustic cues (their cue weighting). These patterns appear to change with increased linguistic experience. Previous speech perception research has found a positive correlation between more analytical cue weighting strategies and the ability to consciously think about and manipulate segment-sized units (phonemic awareness). That research did not, however, aim to address whether the relation is in any way causal or, if so, then in which direction possible causality might move. Causality in this relation could move in 1 of 2 ways: Either phonemic awareness development could impact on cue weighting strategies or changes in cue weighting could allow for the later development of phonemic awareness. The aim of this study was to follow the development of these 2 processes longitudinally to determine which of the above 2 possibilities was more likely. Five-year-old children were tested 3 times in 7 months on their cue weighting strategies for a /so/-/[symbol in text]o/ contrast, in which the 2 cues manipulated were the frequency of fricative spectrum and the frequency of vowel-onset formant transitions. The children were also tested at the same time on their phoneme segmentation and phoneme blending skills. Results showed that phonemic awareness skills tended to improve before cue weighting changed and that early phonemic awareness ability predicted later cue weighting strategies. These results suggest that the development of metaphonemic awareness may play some role in changes in cue weighting.

Adult↗

[Speech perception. The basis for speech audiometry examinations].

Speech recognition measurements are widely used in clinical applications. Usually, either a speech recognition threshold is determined, being defined as the speech level or signal-to-noise ratio that enables the patient to understand 50% of the target words or target sentences. In other cases, a discrimination curve, describing speech recognition as function of the presentation level, is recorded. For the interpretation of these findings, it is important to keep in mind that speech perception is affected by many interacting sensory, perceptual and cognitive processes.Here, our current knowledge of these processes is outlined and theories that model speech perception in general are briefly reviewed.

Audiometry, Speech↗

Auditory-visual speech perception and synchrony detection for speech and nonspeech signals.

Previous research has identified a "synchrony window" of several hundred milliseconds over which auditory-visual (AV) asynchronies are not reliably perceived. Individual variability in the size of this AV synchrony window has been linked with variability in AV speech perception measures, but it was not clear whether AV speech perception measures are related to synchrony detection for speech only or for both speech and nonspeech signals. An experiment was conducted to investigate the relationship between measures of AV speech perception and AV synchrony detection for speech and nonspeech signals. Variability in AV synchrony detection for both speech and nonspeech signals was found to be related to variability in measures of auditory-only (A-only) and AV speech perception, suggesting that temporal processing for both speech and nonspeech signals must be taken into account in explaining variability in A-only and multisensory speech perception.

Acoustic Stimulation↗

The dynamic nature of speech perception.

The speech perception system must be flexible in responding to the variability in speech sounds caused by differences among speakers and by language change over the lifespan of the listener. Indeed, listeners use lexical knowledge phonetic to retune perception of novel speech (Norris, McQueen, & Cutler, 2003). In categorization that study, Dutch listeners made lexical decisions to spoken stimuli, including words with an ambiguous fricative (between [f] and [s]), in either [f]- or speech [s]-biased lexical contexts. In a subsequent categorization test, the former perception group of listeners identified more sounds on an [epsilonf]-[epsilons] continuum as [f] than the latter group. In the present experiment, listeners received the same exposure and test stimuli, but did not make lexical decisions to the exposure items. Instead, they counted them. Categorization results were indistinguishable from those obtained earlier. These adjustments in fricative perception therefore do not depend on explicit judgments during exposure. This learning effect thus reflects automatic retuning of the interpretation of acoustic-phonetic information.

Humans↗

Studying speech perception in adolescent school-age children by utilizing primary color perception.

Since both speech and color have perceptual ramifications in language, the present study was developed to study speech perception through color perception. With the recent advances in perception generally and color perception specifically, this nontraditional approach to studying speech perception appeared reasonable. The 12 consonants utilized in this study generated 144 pairs of nonsense CV-syllables. The consonant ensemble was selected because it accommodated a maximum number of phonological features with a minimum number of phonemes. The 46 subjects, who were in junior high school with an age range of 11 through 14 years, responded to each of the stimulus pairs by assigning (associating) to it one of the six primary colors. Because of the perceptual orderliness associated with subjects' judgments, the results indicated that color can be used to study speech perception. Specific findings included the retrieval of sameness, fronting (or place), and voicing.

Adolescent↗

Speech perception without traditional speech cues.

A three-tone sinusoidal replica of a naturally produced utterance was identified by listeners, despite the readily apparent unnatural speech quality of the signal. The time-varying properties of these highly artificial acoustic signals are apparently sufficient to support perception of the linguistic message in the absence of traditional acoustic cues for phonetic segments.

Auditory Perception↗

Cognitive factors and cochlear implants: some thoughts on perception, learning, and memory in speech perception.

Over the past few years, there has been increased interest in studying some of the cognitive factors that affect speech perception performance of cochlear implant patients. In this paper, I provide a brief theoretical overview of the fundamental assumptions of the information-processing approach to cognition and discuss the role of perception, learning, and memory in speech perception and spoken language processing. The information-processing framework provides researchers and clinicians with a new way to understand the time-course of perceptual and cognitive development and the relations between perception and production of spoken language. Directions for future research using this approach are discussed including the study of individual differences, predicting success with a cochlear implant from a set of cognitive measures of performance and developing new intervention strategies.

Child↗

Multisensory speech perception of young children with profound hearing loss.

The contribution of a two-channel vibrotactile aid (Trill VTA 2/3, AVR Communications LTD) to the audiovisual perception of speech was evaluated in four young children with profound hearing loss using words and speech pattern contrasts. An intensive, hierarchical, and systematic training program was provided. The results show that the addition of the tactile (T) modality to the auditory and visual (A+V) modalities enhanced speech perception performance significantly on all tests. Specifically, at the end of the training sessions, the tactile supplementation increased word recognition scores in a 44-word, closed-set task by 12 percentage points; detection of consonant in final position by 50 percentage points; detection of sibilant in final position by 30 percentage points; and detection of voicing in final position by 25 percentage points. Significant learning over time was evident for all test materials, in all modalities. As expected, fastest learning (i.e., smallest time constants) was found for the AVT condition. The results of this study provide further evidence that sensory information provided by the tactile modality can enhance speech perception in young children.

Child, Preschool↗

Word recognition in a foreign language: a study of speech perception.

Models of speech perception have stressed the importance of investigating recognition of words in fluent speech. The effects of word length and the initial phonemes of words on the speech perception of foreign language learners were investigated. English-speaking subjects were asked to listen for target words in repeated presentations of a prose passage read in French by a native speaker. The four target words were either one or four syllables in length and began with either an initial stop or fricative consonant. Each of the four words was substituted 60 times in identical sentence contexts in place of nouns deleted from the original story. The results indicated that four-syllable words were more easily detected than one-syllable words. Contrary to expectation, stop-initial words were not more accurately detected than fricative-initial words. Based on these findings additional considerations that seem needed in order to apply current models of word recognition to naive listeners are discussed.

Humans↗

Speech perception with the ACE and the SPEAK speech coding strategies for children implanted with the Nucleus cochlear implant.

OBJECTIVE: The aim of this study is to determine whether implanted children using the ACE speech coding strategy demonstrate superior performances compared to implanted children using the SPEAK speech coding strategy over time. METHODS: Cochlear implanted children with prelinguistic sensorineural bilateral deafness of profound degree, using either the ACE or SPEAK coding strategy, were evaluated and compared. Both groups of children used one of the speech coding strategies continuously from the initial programming session and for a period of 2 years post-switch-on. One group comprised children who were retrospectively implanted and had received the SPEAK speech coding strategy (n=32) and the second group consisted of prospectively implanted children who received the ACE speech coding strategy (n=26). Both populations were homogenous as far as age of implantation, degree of hearing loss, anatomy of the cochlea, depth of electrode insertion, and educational and rehabilitative support provided. Children were assessed at 6, 12 and 24 months post switch-on via pure-tone audiometry and for speech perception tests. Children using the ACE speech coding strategy were additionally evaluated using the MAIS and MUSS language scales. RESULTS: Satisfactory benefits in speech perception were demonstrated by both groups of implanted children. No significant difference between the mean pure tone thresholds was observed postoperatively between the groups. Two years post switch-on the group using the ACE speech coding strategy demonstrated superior results for vowel discrimination in comparison to children using the SPEAK coding strategy. No significant difference was observed between the groups for performance on discrimination of syllable patterns (ESP) or for disyllablic word recognition tests. Additionally, the group of ACE users demonstrated maximum performance on MAIS and MUSS scales, 2 years post switch-on. CONCLUSIONS: The results clearly demonstrate significant benefit of cochlear implantation in prelinguistically deafened children for speech perception ability when using either the SPEAK or ACE speech coding strategies. Children using the ACE speech coding strategy demonstrate more rapid progress in improved speech perception ability initially, however 2 years post switch-on, no significant difference in performance on open-set speech recognition tests can be noted irrespective of the strategy in use.

Audiometry, Pure-Tone↗

La langue et les Lèvres: cross-language influences on bimodal speech perception.

Previous research in speech perception has yielded two sets of findings which are brought together in the present study. First, it has been shown that normal hearing listeners use visible as well as acoustical information when processing speech. Second, it has been shown that there is an effect of specific language experience on speech perception such that adults often have difficulty identifying and discriminating non-native phones. The present investigation was designed to extend and combine these two sets of findings. Two studies were conducted using six consonant-vowel syllables (/ba/, /va/, /alpha a/, /da/, /3a/, and /ga/ five of which occur in French and English, and one (the interdental fricative /alpha a/) which occurs only in English. In Experiment 1, an effect of specific linguistic experience was evident for the auditory identification of the non-native interdental stimulus by French-speakers. In Experiment 2, it was shown that the effect of specific language experience extends to the perception of the visible information in speech. These findings are discussed in terms of their implications for our understanding of cross-language processes in speech perception and for our understanding of the development of bimodal speech perception.

Adult↗

Speech perception in older adults: the importance of speech-specific cognitive abilities.

OBJECTIVE: To provide a critical evaluation of studies examining the contribution of changes in language-specific cognitive abilities to the speech perception difficulties of older adults. DESIGN: A review of the literature on aging and speech perception. CONCLUSIONS: The research considered in the present review suggests that age-related changes in absolute sensitivity is the principal factor affecting older listeners' speech perception in quiet. However, under less favorable listening conditions, changes in a number of speech-specific cognitive abilities can also affect spoken language processing in older people. Clinically, these findings suggest that hearing aids, which have been the traditional treatment for improving speech perception in older adults, are likely to offer considerable benefit in quiet listening situations because the amplification they provide can serve to compensate for age-related hearing losses. However, such devices may be less beneficial in more natural environments, (e.g., noisy backgrounds, multiple talkers, reverberant rooms) because they are less effective for improving speech perception difficulties that result from age-related cognitive declines. It is suggested that an integrative approach to designing test batteries that can assess both sensory and cognitive abilities needed for processing spoken language offers the most promising approach for developing therapeutic interventions to improve speech perception in older adults.

Aged↗

Are there interactive processes in speech perception?

Lexical information facilitates speech perception, especially when sounds are ambiguous or degraded. The interactive approach to understanding this effect posits that this facilitation is accomplished through bi-directional flow of information, allowing lexical knowledge to influence pre-lexical processes. Alternative autonomous theories posit feed-forward processing with lexical influence restricted to post-perceptual decision processes. We review evidence supporting the prediction of interactive models that lexical influences can affect pre-lexical mechanisms, triggering compensation, adaptation and retuning of phonological processes generally taken to be pre-lexical. We argue that these and other findings point to interactive processing as a fundamental principle for perception of speech and other modalities.

Humans↗

Speech perception in children using the advanced Speak speech-processing strategy.

The Speak speech-processing strategy, developed by the University of Melbourne and commercialized by Cochlear Pty Limited for use in the new Spectra 22 speech processor, has been shown to provide improved speech perception for adults in both quiet and noisy situations. The present study evaluated the ability of children experienced in the use of the Multipeak (Mpeak) speech-processing strategy (implemented in the Nucleus Minisystem-22 cochlear implant) to adapt to and benefit from the advanced Speak speech-processing strategy (implemented in the Nucleus Spectra 22 speech processor). Twelve children were assessed using Mpeak and Speak over a period of 8 months. All of the children had over 1 year's previous experience with Mpeak, and all were able to score significantly on open-set word and sentence tests using the cochlear implant alone. Children were assessed with both live-voice and recorded speech materials, including Consonant-Nucleus-Consonant monosyllabic words and Speech Intelligibility Test sentences. Assessments were made in both quiet and in noise. Assessments were made at 3-week intervals to investigate the ability of the children to adapt to the new speech-processing strategy. For most of the children, a significant advantage was evident when using the Speak strategy as compared with Mpeak. For 4 of the children, there was no decrement in speech perception scores immediately following fitting with Speak. Eight of the children showed a small (10% to 20%) decrement in speech perception scores for between 3 and 6 weeks following the changeover to Speak. After 24 weeks' experience with Speak, 11 of the children had shown a steady increase in speech perception scores, with final Speak scores higher than for Mpeak. Only 1 child showed a significant decrement in speech perception with Speak, which did not recover to original Mpeak levels.

Adolescent↗

Speech perception in noise with implant and hearing aid.

OBJECTIVE: To compare the perception of speech in quiet and in noise by adults using a cochlear implant on its own or a cochlear implant and hearing aid together. STUDY DESIGN: Repeated measures. SETTING: Laboratory study using subjects' own speech processors. PATIENTS: Two groups of cochlear implant users (Australian and American) with some residual hearing in the non-implanted ear (pure tone average thresholds at 500 Hz, 1 kHz, and 2 kHz of 75-112 dB HL). INTERVENTION(S): Conventional hearing aid and cochlear implant in opposite ears. MAIN OUTCOME MEASURES: Speech perception was evaluated using recorded lists of CUNY sentences and lists of CNC words in quiet and in background noise. RESULTS: Speech scores were significantly higher with implant and hearing aid together compared to implant alone. The binaural advantage was greater in background noise than it was in quiet for CUNY sentences in the American listeners. CONCLUSION: Severely-to-profoundly hearing impaired adults may benefit from combined fitting of implants and conventional hearing aids in opposite ears.

Adult↗