PubMed · 15215465
Complexity: an internet resource for analysis of DNA sequence complexity.
Abstract
The search for DNA regions with low complexity is one of the pivotal tasks of modern structural analysis of complete genomes. The low complexity may be preconditioned by strong inequality in nucleotide content (biased composition), by tandem or dispersed repeats or by palindrome-hairpin structures, as well as by a combination of all these factors. Several numerical measures of textual complexity, including combinatorial and linguistic ones, together with complexity estimation using a modified Lempel-Ziv algorithm, have been implemented in a software tool called 'Complexity' (http://wwwmgs.bionet.nsc.ru/mgs/programs/low_complexity/). The software enables a user to search for low-complexity regions in long sequences, e.g. complete bacterial genomes or eukaryotic chromosomes. In addition, it estimates the complexity of groups of aligned sequences.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Y L Orlov, V N Potapov. 2004-07-01. Complexity: an internet resource for analysis of DNA sequence complexity.. https://doi.org/10.1093/nar%2Fgkh466
Cite the original work for its findings. Save a collection to share your selection of sources.