PubMed · 11604875
Protein structural domain parsing by consensus reasoning over multiple knowledge sources and methods.
Abstract
Domain parsing, or the detection of signals of protein structural domains from sequence data, is a complex and difficult problem. If carried out reliably it would be a powerful interpretive and predictive tool for genomic and proteomic studies. We report on a novel approach to domain parsing using consensus techniques based on Hidden Markov Models (HMMs) and BLAST searches built from a training set of 1471 continuous structural domains from the Dali Domain Dictionary (DDD). Validation on an independent test sample of family-matched structural domain sequences from the Scop database yields a consensus prediction performance rate of 75.5%, well above the 58% obtained by simple agreement of methods.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
C A Kulikowski, I Muchnik, H J Yun, A A Dayanik, D Zhang, Y Song, G T Montelione. 2001. Protein structural domain parsing by consensus reasoning over multiple knowledge sources and methods.. https://pubmed.ncbi.nlm.nih.gov/11604875/
Cite the original work for its findings. Save a collection to share your selection of sources.