PubMed · 15838139
High similarity sequence comparison in clustering large sequence databases.
Abstract
We present a fast algorithm for sequence clustering and searching which works with large sequence databases. It uses a strictly defined similarity measure. The algorithm is faster than conventional EST clustering approaches because its complexity is directly related to the number of subwords shared by the sequences. Furthermore, the algorithm also works with proteic sequences and large sequences like entire chromosomes. We present a theoretical study of our approach and provide experimental results.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Lorie Dudoignon, Eric Glemet, Hendrik Cornelis Heus, Mathieu Raffinot. 2002. High similarity sequence comparison in clustering large sequence databases.. https://pubmed.ncbi.nlm.nih.gov/15838139/
Cite the original work for its findings. Save a collection to share your selection of sources.