PubMed Health⌕ Search

PubMed · 7479810

Training and search methods for speech recognition.

Abstract

Speech recognition involves three processes: extraction of acoustic indices from the speech signal, estimation of the probability that the observed index string was caused by a hypothesized utterance segment, and determination of the recognized utterance via a search among hypothesized alternatives. This paper is not concerned with the first process. Estimation of the probability of an index string involves a model of index production by any given utterance segment (e.g., a word). Hidden Markov models (HMMs) are used for this purpose [Makhoul, J. & Schwartz, R. (1995) Proc. Natl. Acad. Sci. USA 92, 9956-9963]. Their parameters are state transition probabilities and output probability distributions associated with the transitions. The Baum algorithm that obtains the values of these parameters from speech data via their successive reestimation will be described in this paper. The recognizer wishes to find the most probable utterance that could have caused the observed acoustic index string. That probability is the product of two factors: the probability that the utterance will produce the string and the probability that the speaker will wish to produce the utterance (the language model probability). Even if the vocabulary size is moderate, it is impossible to search for the utterance exhaustively. One practical algorithm is described [Viterbi, A. J. (1967) IEEE Trans. Inf. Theory IT-13, 260-267] that, given the index string, has a high likelihood of finding the most probable utterance.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

F Jelinek. 1995-10-24. Training and search methods for speech recognition.. https://doi.org/10.1073/pnas.92.22.9964

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Information-sharing among couples considering multifetal pregnancy reduction.

OBJECTIVE: To determine the information-sharing strategies of couples considering fetal reduction, and the impact of these strategies on the chances of encountering hostility in their social networks. DESIGN: Cross-sectional design of semistructured qualitative interviews, coded with respect to sharing strategies and level of personally directed hostility encountered. SETTING: Multiple Pregnancy Management Program, Comprehensive Genetics, New York, New York. PATIENT(S) AND INTERVENTION(S): Fifty women and their partners who were making a first visit to our maternal-fetal management facility, in order to consider the possibility of multifetal reduction as a pregnancy-management strategy. MAIN OUTCOME MEASURE(S): Development of information-sharing strategies, and the chances of encountering personally directed hostility regarding multifetal reduction associated with more and less selective strategies. RESULT(S): Four information-sharing strategies emerged from the analysis. Two of these strategies were relatively open (extended network, and both parents). Two other strategies were relatively selective (qualified family and friends, and defended relationship). The selective strategies were significantly less likely to encounter to encounter personally directed hostility (odds ratio, 3.88; 95% confidence intervals, 0.87-17.30). CONCLUSION(S): Selective sharing of information for couples considering multifetal prgnancy reduction is a potentially useful strategy for moderating potentially stressful relationships in their social networks. Clinics should find a way of integrating the discussion of selective sharing into their clinic's cultural repertoire of patient-support services.

Communication↗