PubMed HealthSearch

PubMed · 8290362

Feature expressions: creating and manipulating sequence datasets.

Abstract

Annotation of features, such as introns, exons and protein coding regions in GenBank/EMBL/DDBJ entries is now standardized through use of the Features Table (FT) language. The essence of the FT language is described by the relation 'expression-->sequence', meaning that each FT expression evaluates to a sequence. For example, the expression M74750:1..50 evaluates to the first 50 bases of the sequence with accession number M74750. Because FT is intrinsic to the database definition, it can serve as a software- and platform-independent lingua franca for sequence manipulation. The XYLEM package makes it possible to create and manipulate sequence datasets using FT expressions. FEATURES is a program that resolves FT expressions into their corresponding sequences. Annotated features can be retrieved either by feature key or by expression. Even unannotated portions of a sequence can be retrieved by user-generated FT expressions. Applications of the FT language include retrieval of subsequences from large sequence entries, generation of chromosome models or artificial DNA constructs, and representation of restriction maps or mutants.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

B Fristensky. 1993-12-25. Feature expressions: creating and manipulating sequence datasets.. https://doi.org/10.1093/nar%2F21.25.5997

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

miR-503-3p promotes epithelial-mesenchymal transition in breast cancer by directly targeting SMAD2 and E-cadherin.

Although progress in clinical and basic research has significantly increased our understanding of breast cancer, little is known about the molecular mechanism underlying breast cancer metastasis. Identification of effective therapeutic targets to prevent breast cancer metastasis is urgently needed. The function of miR-503-3p has been investigated in other cancers, but its role in breast cancer remains undefined. Here, we found that miR-503-3p was overexpressed in breast cancer tissue and plasma compared with adjacent normal breast tissue and with plasma from healthy individuals. Moreover, we identified miR-503-3p to be an oncogene of breast cancer cell proliferation, migration and invasion. Upregulation of miR-503-3p in breast cancer cells inhibited expression of epithelial-mesenchymal transition (EMT)-related protein SMAD2 and the epithelial marker protein E-cadherin by directly binding to their mRNA 3' untranslated region, whereas increased expression of mesenchymal marker proteins, including vimentin and N-cadherin. Taken together, our findings support a critical role for miR-503-3p in induction of breast cancer EMT and suggest that plasma miR-503-3p may be a useful diagnostic biomarker for breast cancer.

Base Sequence

Identification and characterization of Prp45p and Prp46p, essential pre-mRNA splicing factors.

Through exhaustive two-hybrid screens using a budding yeast genomic library, and starting with the splicing factor and DEAH-box RNA helicase Prp22p as bait, we identified yeast Prp45p and Prp46p. We show that as well as interacting in two-hybrid screens, Prp45p and Prp46p interact with each other in vitro. We demonstrate that Prp45p and Prp46p are spliceosome associated throughout the splicing process and both are essential for pre-mRNA splicing. Under nonsplicing conditions they also associate in coprecipitation assays with low levels of the U2, U5, and U6 snRNAs that may indicate their presence in endogenous activated spliceosomes or in a postsplicing snRNP complex.

Base Sequence

Determination of human DNA polymerase utilization for the repair of a model ionizing radiation-induced DNA strand break lesion in a defined vector substrate.

Human DNA polymerase and DNA ligase utilization for the repair of a major class of ionizing radiation-induced DNA lesion [DNA single-strand breaks containing 3'-phosphoglycolate (3'-PG)] was examined using a novel, chemically defined vector substrate containing a single, site-specific 3'-PG single-strand break lesion. In addition, the major human AP endonuclease, HAP1 (also known as APE1, APEX, Ref-1), was tested to determine if it was involved in initiating repair of 3'-PG-containing single-strand break lesions. DNA polymerase beta was found to be the primary polymerase responsible for nucleotide incorporation at the lesion site following excision of the 3'-PG blocking group. However, DNA polymerase delta/straightepsilon was also capable of nucleotide incorporation at the lesion site following 3'-PG excision. In addition, repair reactions catalyzed by DNA polymerase beta were found to be most effective in the presence of DNA ligase III, while those catalyzed by DNA polymerase delta/straightepsilon appeared to be more effective in the presence of DNA ligase I. Also, it was demonstrated that the repair initiating 3'-PG excision reaction was not dependent upon HAP1 activity, as judged by inhibition of HAP1 with neutralizing HAP1-specific polyclonal antibody.

Base Sequence