PubMed Health⌕ Search

PubMed · 11928468

Scoring pairwise genomic sequence alignments.

Abstract

The parameters by which alignments are scored can strongly affect sensitivity and specificity of alignment procedures. While appropriate parameter choices are well understood for protein alignments, much less is known for genomic DNA sequences. We describe a straightforward approach to scoring nucleotide substitutions in genomic sequence alignments, especially human-mouse comparisons. Scores are obtained from relative frequencies of aligned nucleotides observed in alignments of non-coding, non-repetitive genomic regions, and can be theoretically motivated through substitution models. Additional accuracy can be attained by down-weighting alignments characterized by low compositional complexity. We also describe an evaluation protocol that is relevant when alignments are intended to identify all and only the orthologous positions. One particular scoring matrix, called HOXD70, has proven to be generally effective for human-mouse comparisons, and has been used by the PipMaker server since July, 2000. We discuss but leave open the problem of effectively scoring regions of strongly biased nucleotide composition, such as low G + C content.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

F Chiaromonte, V B Yap, W Miller. 2002. Scoring pairwise genomic sequence alignments.. https://doi.org/10.1142/9789812799623_0012

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Application of mutated miR-206 target sites enables skeletal muscle-specific silencing of transgene expression of cardiotropic AAV9 vectors.

Insertion of completely complementary microRNA (miR) target sites (miRTS) into a transgene has been shown to be a valuable approach to specifically repress transgene expression in non-targeted tissues. miR-122TS have been successfully used to silence transgene expression in the liver following systemic application of cardiotropic adeno-associated virus (AAV) 9 vectors. For miR-206-mediated skeletal muscle-specific silencing of miR-206TS-bearing AAV9 vectors, however, we found this approach failed due to the expression of another member (miR-1) of the same miR family in heart tissue, the intended target. We introduced single-nucleotide substitutions into the miR-206TS and searched for those which prevented miR-1-mediated cardiac repression. Several mutated miR-206TS (m206TS), in particular m206TS-3G, were resistant to miR-1, but remained fully sensitive to miR-206. All these variants had mismatches in the seed region of the miR/m206TS duplex in common. Furthermore, we found that some m206TS, containing mismatches within the seed region or within the 3' portion of the miR-206, even enhanced the miR-206- mediated transgene repression. In vivo expression of m206TS-3G- and miR-122TS-containing transgene of systemically applied AAV9 vectors was strongly repressed in both skeletal muscle and the liver but remained high in the heart. Thus, site-directed mutagenesis of miRTS provides a new strategy to differentiate transgene de-targeting of related miRs.

Base Pairing↗

Sequence context-dependent replication of DNA templates containing UV-induced lesions by human DNA polymerase iota.

Humans possess four Y-family polymerases: pols eta, iota, kappa and the Rev1 protein. The pivotal role that pol eta plays in protecting us from UV-induced skin cancers is unquestioned given that mutations in the POLH gene (encoding pol eta), lead to the sunlight-sensitive and cancer-prone xeroderma pigmentosum variant phenotype. The roles that pols iota, kappa and Rev1 play in the tolerance of UV-induced DNA damage is, however, much less clear. For example, in vitro studies in which the ability of pol iota to bypass UV-induced cyclobutane pyrimidine dimers (CPDs) or 6-4 pyrimidine-pyrimidone (6-4PP) lesions has been assayed, are somewhat varied with results ranging from limited misinsertion opposite CPDs to complete lesion bypass. We have tested the hypothesis that such discrepancies might have arisen from different assay conditions and local sequence contexts surrounding each UV-photoproduct and find that pol iota can facilitate significant levels of unassisted highly error-prone bypass of a T-T CPD, particularly when the lesion is located in a 3'-A[T-T]A-5' template sequence context and the reaction buffer contains no KCl. When encountering a T-T 6-4PP dimer under the same assay conditions, pol iota efficiently and accurately inserts the correct base, A, opposite the 3'T of the 6-4PP by factors of approximately 10(2) over the incorporation of incorrect nucleotides, while incorporation opposite the 5'T is highly mutagenic. Pol kappa has been proposed to function in the bypass of UV-induced lesions by helping extend primers terminated opposite CPDs. However, we find no evidence that the combined actions of pol iota and pol kappa result in a significant increase in bypass of T-T CPDs when compared to pol iota alone. Our data suggest that under certain conditions and sequence contexts, pol iota can bypass T-T CPDs unassisted and can efficiently incorporate one or more bases opposite a T-T 6-4PP. Such biochemical activities may, therefore, be of biological significance especially in XP-V cells lacking the primary T-T CPD bypassing enzyme, pol eta.

Base Pairing↗