PubMed HealthSearch

SEARCH · PubMed Health

Results for “Speech Production Measurement”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 recordsLinked to original sources

Vocal roughness and jitter characteristics of vowels produced by esophageal speakers.

Audiotape recordings of sustained vowels produced by nine esophageal speakers were subjected to acoustic and perceptual analysis. Results indicated that (1) the magnitude of vocal jitter present in the vowels was substantially larger than that observed in normal speakers and speakers with laryngeal/vocal disturbance, (2) listeners could reliably rate the severity of vocal roughness in the vowels, (3) voices of esophageal speakers were characterized by varying degrees of vocal roughness, and (4) mean fundamental frequency, mean jitter, or jitter ratio measures did not serve as useful predictors of the perceived severity of vocal roughness. These findings are interpreted to suggest that the mechanism esophageal speakers employ to regulate fundamental frequency is substantially different from that employed by normal speakers and that the identity of physical variables underlying the perception of roughness severity in naturally produced human speech is not well understood.

Acoustics

Treatment of elective mute behavior in two developmentally delayed children using modeling and contingency management.

Most classification schemes differentiate elective mutism from language problems seen in the developmentally delayed population. Two preschool developmentally delayed children were treated for speech reluctance using modeling and contingency management. Employing a multiple baseline across therapists, it was found that these treatment components were effective in increasing frequency of labeling behavior in both children. Results were maintained at follow-up. Generalization to new words and to spontaneous speech were also noted, and suggest that characteristics of elective mutism in this population may be similar to what is found in the general population.

Child Language

Esophageal speaker articulation of /s,z/: a dynamic palatometric assessment.

Esophageal talker linguapalatal contact patterns and durations during /s/ and /z/ productions were examined using dynamic palatometry instrumentation. It was found that sibilant groove narrowing is a physiologic compensation for a reduced air supply in esophageal speech. The place of esophageal /s, z/ articulation was on the anterior portion of the alveolar ridge as seen in normal speakers. Average medial groove width for esophageal /s/ was narrower than the 5-7-mm groove characteristic of normal speakers. Groove widths averaged 3 mm for /s/ and 4 mm for /z/. Systematic changes in groove widths across speech sounds, syllable position, and vowel context were also observed. Use of a narrower lingual groove was interpreted as a significant articulatory maneuver to meter out a limited intraoral air supply and effect more normal fricative durations.

Articulation Disorders

[Voice rehabilitation after laryngectomy with voice prostheses. Results of a prospective follow-up study].

Although the results of surgical rehabilitation by means of voice prostheses are on the average better than rehabilitation via oesophageal speech, the tracheoesophageal puncture (TEP)-technique has so far not been widely used in Germany. The majority of hospitals still prefer the "traditional" method of voice rehabilitation using oesophageal speech. The present prospective study was undertaken to compare the results of postlaryngectomy vocal rehabilitation, if patients were offered the surgical voice rehabilitation via voice prosthesis as an alternative to oesophageal speech. Taking into account all the patients who underwent laryngectomy from 1989 until 1990 in Tübingen, primary surgical voice rehabilitation was performed in 44 out of 54 patients (81.5%). Interestingly enough, 34 patients who underwent laryngectomy were able to perform communication via the telephone on the day of their discharge. Moreover, one-third of the laryngectomised patients showed a significant increase in speech intelligibility within the first six months after laryngectomy. 36 patients with laryngectomy were able to attain proficiency 6 months after surgery. In 12 patients the prosthesis had to be removed, since either phonation was impossible or patients successfully learned and preferred oesophageal speech. In conclusion, independent of the method of voice rehabilitation (prosthesis, electrolarynx, oesophageal speech), our results support the hypothesis that a voice rehabilitation regimen will yield a higher rehabilitation rate of patients if rehabilitation via surgical voice is offered as an alternative to learning the oesophageal voice. Therefore, it seems to be advisable that patients are allowed to have the choice between surgical rehabilitation and oesophageal speech restoration.(ABSTRACT TRUNCATED AT 250 WORDS)

Female

Perceptual comparison of neoglottal, oesophageal and normal speech.

The sex, normality, intelligibility, rate, rhythm and intonation of 44 neoglottal, oesophageal and normal speakers have been judged by a panel of 10 trained listeners. It was found that the sex of the alaryngeal speakers was perceived correctly much less reliably than that of the normal speakers. Neoglottal speakers were rated more normal and intelligible than oesophageal speakers. The speaking rate of neoglottal speakers was judged not to be significantly different from that of normal speakers. Neoglottal speakers were considered more fluent and to have better intonation than the other alaryngeal speakers. Thus neoglottal speakers were found, on average, to be as good as or better than good oesophageal speakers in each of the respects of which judgements were made.

Adult

Acoustic invariance in speech production: evidence from measurements of the spectral characteristics of stop consonants.

On the basis of theoretical considerations and the results of experiments with synthetic consonant-vowel syllables, it has been hypothesized that the short-time spectrum sampled at the onset of a stop consonant should exhibit gross properties that uniquely specify the consonantal place of articulation independent of the following vowel. The aim of this paper is to test this hypothesis by measuring the spectrum sampled at the onsets and offsets of a large number of consonant-vowel (CV) and vowel-consonant (VC) syllables containing both voiced and voiceless stops produced by several speakers. Templates were devised in an attempt to capture three classes of spectral shapes: diffuse-rising, diffuse-falling, and compact, corresponding to alveolar, labial, and velar consonants, respectively. Spectra were derived from the utterances by sampling at the consonantal release of CV syllables and at the implosion and burst release of VC syllables, and these spectra (smoothed by a linear prediction algorithm) were matched against the templates. It was found that about 85% of the spectra at initial consonant release and at final burst release were correctly classified by the templates, although there was some variability across vowel contexts. The spectra sampled at the implosion were not consistently classified. A preliminary examination of spectra sampled at the release of nasal consonants in CV syllables showed a somewhat lower accuracy of classification by the same templates. Overall, the results support an hypothesis that, in natural speech, the acoustic characteristics of stop consonants, specified in terms of the gross spectral shape sampled at the discontinuity in the acoustic signal, show invariant properties independent of the adjacent vowel or of the voicing characteristics of the consonant. The implication is that the auditory system is endowed with detectors that are sensitive to these kinds of gross spectral shapes, and that the existence of these detectors helps the infant to organize the sounds of speech into their natural classes.

Humans

Toward measuring how well hearing-impaired children speak.

Average intelligibility scores for a group of 37 hearing-impaired and two normally hearing adolescents were determined by 50 normal listeners and were compared with nine acoustically measured speech variables. These nine variables included measurements of consonant production, vowel production, and prosody. Regression analysis of the variables showed that three of the speech variables bore a multiple correlation of 0.85 with measured intelligibility scores. Two variables alone, the mean voice-onset-time difference between /t/ and /d/ and the mean second-formant difference between /i/ and /c/, accounted for about 70% of the variance in the intelligibility scores. To cross-validate the reliability of these correlations, intelligibility scores were subsequently predicted for another group of 30 hearing-impaired adolescents and then compared with intelligibility scores as determined by another group of normal listeners. For this second group, the correlation between measured intelligibility scores and predicted scores was 0.86, which indicates that the reliability of the predicting variables is high. Five of the nine variables correlated more highly with measured speech intelligibility than did pure-tone audiometric thresholds. The average speech intelligibility of all 67 hearing-impaired subjects was 76%.

Acoustics

Comparing Traditional Motor Speech Practice to Contextualized Speech Practice in Preschoolers With Childhood Apraxia of Speech.

PURPOSE: The aim of this study was to compare retention of real-word targets across practice conditions (contextualized vs. motor-only) within a modified integral stimulation treatment for preschoolers with childhood apraxia of speech (CAS). METHOD: A single-subject experimental design with alternating treatments was used with matched target sets randomly assigned to contextualized practice, motor-only practice, or no treatment. Three preschoolers with CAS completed 18 therapy sessions, each consisting of two 25-min blocks: one contextualized practice and one motor-only practice. Order of practice was randomized each visit. Changes in percent phonemes correct (PPC) and lexical stress accuracy, derived from blinded transcription, were explored with visual analysis and effect sizes (standardized mean difference, d statistic). RESULTS: Meaningful improvements (d > 1) were observed in PPC across words treated in contextualized practice for all three children immediately posttreatment and for two of three children at the 1-month follow-up. Meaningful improvements in the motor-only condition were observed in two of three children immediately posttreatment and at follow-up. No meaningful changes were observed in lexical stress across any conditions in any participant. CONCLUSIONS: This study provides preliminary support for the feasibility of a modified integral stimulation therapy that incorporates elements of linguistically grounded therapies (linguistic retrieval, recasts, expansions) that may facilitate target retention in some preschoolers with CAS. However, other elements should be explored in conjunction with integral stimulation to maximize clinical outcomes. SUPPLEMENTAL MATERIAL: https://doi.org/10.23641/asha.33228981.

Humans

Digital profiling of dysarthria in late- and early-onset Parkinson's disease.

BackgroundDigital speech analysis affords robust markers of Parkinson's disease (PD). However, most studies target late-onset PD (LOPD), neglecting early-onset PD (EOPD) -an increasingly prevalent subtype. This proof-of-concept study tackles such gap.MethodsWe used machine learning to discriminate persons with EOPD (with symptom onset before age 50) and LOPD (with symptom onset after age 50) from healthy controls (HCs) through prosodic and articulatory features from natural speech.ResultsMaximal classification between patients and HCs was afforded by combined prosodic and articulation features in LOPD (AUC&#x2009;=&#x2009;0.90) and by articulation alone in EOPD (AUC&#x2009;=&#x2009;0.79), with chance-level discrimination between patient groups (AUC&#x2009;=&#x2009;0.55). Motor severity (MDS-UPDRS-III) scores predicted by these features correlated with actual motor severity scores in both LOPD (r&#x2009;=&#x2009;0.52, p&#x2009;<&#x2009;0.001) and EOPD (r&#x2009;=&#x2009;0.27, p&#x2009;<&#x2009;0.001).ConclusionsDigital speech markers offer markers of PD irrespective of age of onset.Plain language summary titleVoice recordings capture motor symptoms in Parkinson's disease irrespective of age of onset.

Humans

Fundamental frequency perturbation observed in sustained phonation.

The middle segments of sustained phonations of /i/, produced by six male adults at fo's ranging from 98 to 298 Hz, were examined for cycle-to-cycle frequency perturbation. The voice samples were analyzed by a peak-picking fo analysis program (Horii, 1975). The results showed that, between 98 Hz and 210 Hz, mean jitter size decreased as the fo increased, whereas the corresponding jitter ratios remained relatively constant. Above about 210 Hz, mean jitter remained relatively constant and, consequently, the jitter ratio increased as fo increased. The problems of temporal resolution for vocal jitter studies, vowel-dependent jitter magnitude characteristics and various perturbation measures are discussed.

Adult

TONAR calibration: a brief note.

Measurements were made of the frequency response characteristics of the microphone-separator components of TONAR II instrumentation. The results of our calibration studies revealed 1) appreciable non-uniformity in frequency response of the two microphones, 2) a considerable degree of mismatch in frequency response between the microphones and, 3) dynamic interactions among microphone, separator cavity, and talker cavity resonant characteristics. Findings are discussed in terms of their implications regarding the validity of TONAR II based nasalance ratio measures.

Acoustics

Speech production, syntax comprehension, and cognitive deficits in Parkinson's disease.

Speech samples were obtained that were analyzed for voice onset time (VOT) for 40 nondemented English speaking subjects, 20 with mild and 20 with moderate Parkinson's disease. Syntax comprehension and cognitive tests were administered to these subjects in the same test sessions. VOT disruptions for stop consonants in syllable initial position, similar to those noted for Broca's aphasia, occurred for nine subjects. Longer response times and errors in the comprehension of syntax as measured by the Rhode Island Test of Sentence Comprehension (RITLS) also occurred for these subjects. Anovas indicate that the VOT overlap subjects had significantly higher syntax error rates and longer response times on the RITLS than the VOT nonoverlap subjects--F(1, 70) = 12.38, p less than 0.0008; F(1, 70) = 7.70, p less than 0.007, respectively. The correlation between the number of VOT timing errors and the number of syntax errors was significant. (r = 0.6473, p less than 0.01). VOT overlap subjects also had significantly higher error rates in cognitive tasks involving abstraction and the ability to maintain a mental set. Prefrontal cortex, acting through subcortical basal ganglia pathways, is a component of the neural substrate that regulates human speech production, syntactic ability, and certain aspects of cognition. The deterioration of these subcortical pathways may explain similar phenomena in Broca's aphasia. Results are discussed in relation to "modular" theories.

Adult

Speech changes following reimplantation from a single-channel to a multichannel cochlear implant.

The speech of a postlingually deafened preadolescent was recorded and analyzed while a single-electrode cochlear implant (3M/House) was in operation, on two occasions after it failed (1 day and 18 days) and on three occasions after stimulation of a multichannel cochlear implant (Nucleus 22) (1 day, 6 months, and 1 year). Listeners judged 3M/House tokens to be the most normal until the subject had one year's experience with the Nucleus device. Spectrograms showed less aspiration, better formant definition and longer final frication and closure duration post-Nucleus stimulation (6 MO. NUCLEUS and 1 YEAR NUCLEUS) relative to the 3M/House and no auditory feedback conditions. Acoustic measurements after loss of auditory feedback (1 DAY FAIL and 18 DAYS FAIL) indicated a constriction of vowel space. Appropriately higher fundamental frequency for stressed than unstressed syllables, an expansion of vowel space and improvement in some aspects of production of voicing, manner and place of articulation were noted one year post-Nucleus stimulation. Loss of auditory feedback results are related to the literature on the effects of postlingual deafness on speech. Nucleus and 3M/House effects on speech are discussed in terms of speech production studies of single-electrode and multichannel patients.

Child