PubMed HealthSearch

SEARCH · PubMed Health

Results for “Direct RNA sequencing”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5Linked to original sources

Atypical structure of the 23S ribosomal RNA molecule in certain oral bacteria.

Ribosomal RNA (rRNA) isolated from Wolinella recta and seven related bacteria was examined by agarose gel electrophoresis. The 23S rRNA molecule could not be detected in W. recta, Wolinella curva, Bacteroides gracilis, or Bacteroides ureolyticus. In place of the 23S molecule, there were three smaller molecules of approximately 1700, 650, and 600 bases designated 23S alpha, 23S beta, and 23S delta, respectively. An intact 23S rRNA molecule could be isolated from Wolinella succinogenes, Campylobacter concisus, and Campylobacter sputorum. The cleavage sites of the W. recta 23S rRNA molecule were located by direct RNA sequence analysis and were found to be in similar locations, nucleotides 546 and 1180, as cleavage sites described in other prokaryotes. The presence or absence of the 23S rRNA molecule may be a useful marker for these micro-organisms.

Amino Acid Sequence

Translating single-cell RNA sequencing into monocyte direct leukocyte subpopulation-transcript abundance assay ratio-based biomarkers (IFI27/PSAP or IFI27/CTSS) for clinical detection of viral infection.

A rapid method for triaging febrile patients by aetiology (e.g., viral or bacterial infection) using gene expression in peripheral blood (PB) is an intensively researched area. However, gene expression in blood represents a composite sum of gene expression of all the component cell types present in the sample. As a result, numerous genes are measured in most proposed signatures. Herein, we propose a simple ratio-based biomarker (RBB) called direct leukocyte subpopulation-transcript abundance assay (DIRECT LS-TA) that recapitulates gene expressions of a single cell type in PB (i.e., monocytes). Based on single-cell RNA sequencing (scRNAseq) data and bulk expression data, IFI27 and SIGLEC1 are found as interferon-stimulated genes (ISGs) predominantly expressed by monocytes. The DIRECT LS-TA method can use a simple ratio of two genes measured in PB as an RBB to represent the target gene expression in monocytes without the need for monocyte purification. Both scRNAseq and bulk RNA sequencing datasets were used to evaluate the correlation between ISG expression in monocytes and PB, with a particular focus on monocyte expression of IFI27. An iceberg plot of bulk transcriptome data was used to identify genes that were predominantly expressed by monocytes in PB. DIRECT LS-TA RBBs of the three genes (IFI27, IFI44L and SIGLEC1) were evaluated by group-wise comparison, receiver operating characteristic and meta-analysis. In addition, the conventional interferon (IFN) score was evaluated for comparison of diagnostic performance. In viral infection datasets, DIRECT LS-TA of IFI27 (IFI27/PSAP or IFI27/CTSS) was most intensely activated (p value by t test <1e-9) and had the best area under the curve (0.94) among the three potential monocyte ISGs analysed. DIRECT LS-TA SIGLEC1 was also another monocyte biomarker but showed a lower activation (p<9e-5). IFI27/PSAP showed better diagnostic performance than the conventional IFN score. On the other hand, IFI44L was not a predominant monocyte expression gene. DIRECT LS-TA of IFI27 (IFI27/PSAP or IFI27/CTSS) measured in PB was the best biomarker of viral infection and IFN activation among ISGs predominantly expressed by monocytes. It performed even better than the conventional IFN score which required quantification of eight genes. The results suggest that DIRECT LS-TA of IFI27 is a monocyte-informative biomarker which is easy to determine in PB without the need for cell sorting.

Humans

Analysis of formaldehyde-induced Adh mutations in Drosophila by RNA structure mapping and direct sequencing of PCR-amplified genomic DNA.

Two formaldehyde-induced mutations at the Drosophila Adh locus (Adhfn45 and Adhfn46) were analyzed by determining RNA structures at different developmental stages, polymerase chain reaction (PCR) amplification of the affected genomic regions, and direct sequencing of the resulting double-stranded DNA fragments. Adhfn46 adults and larvae accumulate abundant ADH-like distal (adult) and proximal (larval) transcripts that are shorter than transcripts in wild-type flies by a lesion located in the second ADH protein-coding exon. Direct sequencing of the amplified DNA region showed that Adhfn46 contains a 69-bp in-frame deletion that removes 23 amino acids near one border of the second exon. Consistent with these findings, we observed a shorter ADHfn46 protein present at only 3% of wild-type levels. In contrast, Adhfn45 adults and larvae accumulate much smaller amounts of ADH-like distal and proximal transcripts. Both RNAs have an identical aberration in RNA splicing of the 65-base intron sequence. Direct sequencing of the amplified mutated DNA region showed that Adhfn45 contains a 21-bp deletion that removed and rearranged DNA at the 5' splice junction of the 65-bp intron. No ADH cross-reacting material is detected in Adhfn45 flies. Direct-repeat sequences (3-11 bp) are present flanking and within the mutated DNA regions. The patterns of DNA deletion and deletion accompanied by sequence addition at the mutant sites suggest a slipped mispairing mechanism during DNA replication or repair that involves local DNA homology.

Alcohol Dehydrogenase

Analysis of Pteridium ribosomal RNA sequences by rapid direct sequencing.

A total of 864 bases from 5 regions interspersed in the 18S and 26S rRNA molecules from various clones of Pteridium covering the general geographical distribution of the genus was analysed using a rapid rRNA sequencing technique. No base difference has been detected amongst the three major lineages, two of which apparently separated before the breakup of the ancient supercontinent, Pangaea. These regions of the rRNA sequences have thus been conserved for at least 160 million years and are here compared with other eukaryotic, especially plant rRNAs.

Biological Evolution

Energy directed folding of RNA sequences.

A modification of Nussinov's algorithm (1) for (planar) secondary structure generation is described. Our algorithm postpones decisions on matches involving destabiling loops until they prove to be energetically more favourable than more local matches. We present, moreover, an alternative way of representing secondary structures which avoids unwarranted suggestions on higher order neighbourhood, can be automated easily, allows for any amount of annotation of the sequences, makes comparison of alternate foldings easy and is pleasing to the eye. 5S RNA sequences are used to illustrate the methods.

Base Sequence

Genomic and copy-back 3' termini in Sendai virus defective interfering RNA species.

Direct sequencing of nine Sendai virus defective interfering RNA species revealed two kinds of 3'-terminal sequences. Six RNA species had 3' termini identical to the virus genome (negative strand), confirming that internal deletions are a frequent cause of Sendai virus defectiveness. The other three RNA species had 3'-terminal sequences identical to that described as the complement of the 5' terminus of the virus genome (R. A. Lazzarini, J. D. Keene, and M. Schubert, Cell 26:145-154, 1981), indicating that they are of the copy-back type. Extensive homology between these two types of 3' sequences evidently accounts for the ability of the copy-back sequence to function as an initiation signal for viral RNA replication. There may not be a selective advantage of one type of terminus over the other, since one defective interfering strain possessed two RNA species, one of which had the genomic 3' terminus and the other copy-back type.

Base Sequence

Direct nucleotide sequencing using total cellular RNA from dengue-virus-infected mosquito cells.

In recent years, a large amount of nucleotide sequence data for dengue viruses has been published. Most of it was derived by sequencing cDNA synthesized from highly purified genomic viral RNA. This paper presents a simple and rapid method for the isolation of total RNA from mosquito cells infected with dengue viruses. This RNA can be used for direct nucleotide sequencing with specific primers without the need for further purification.

Animals

RNA sequence containing hexanucleotide AAUAAA directs efficient mRNA polyadenylation in vitro.

To determine whether a specific nucleotide sequence is required to direct polyadenylation of a simian virus 40 early pre-mRNA in a soluble HeLa whole-cell lysate, we constructed a series of rearranged and deleted DNA templates, transcribed them in vitro, and determined whether the resultant RNAs could be polyadenylated when incubated in whole-cell lysate. When a 237-base-pair DNA fragment encoding the 3' end of the simian virus 40 early pre-mRNA was transferred to recombinant plasmids encoding RNAs that were not substrates for polyadenylation, the resultant RNAs could now be polyadenylated efficiently. In one case, the chimeric RNA was polyadenylated even more efficiently than was the original simian virus 40 early transcript. Analysis of the RNAs produced from the deletion mutant templates revealed that only RNAs containing at least one copy of the AAUAAA sequence situated near the 3' end and implicated in 3'-end formation and polyadenylation in vivo could be polyadenylated in vitro. Surprisingly, this sequence directed polyadenylation of pre-mRNAs not only when near the RNA 3' end, i.e., 50 nucleotides or less away, but also when the 3' end was situated over 400 nucleotides downstream. Thus, our results show that a polyadenylic acid polymerase activity in HeLa lysates can recognize a specific nucleotide sequence in pre-mRNA and then, in the absence of the nucleolytic cleavage that presumably occurs in vivo, locate the RNA 3' end and use it as a primer for polyadenylic acid synthesis.

Base Sequence

Unusual heterogeneity of leader-mRNA fusion in a murine coronavirus: implications for the mechanism of RNA transcription and recombination.

Coronavirus mRNA transcription was thought to be regulated by the interaction between the leader RNA and the intergenic sequence (IS), probably involving direct RNA-RNA interactions between complementary sequences. In this study, we found that a particular strain of mouse hepatitis virus, JHM2c, which has a deletion of a 9-nucleotide (nt) sequence (UUUAUAAAC) immediately downstream of the leader RNA, transcribed subgenomic mRNA species containing a whole array of heterogeneous leader fusion sites. Using a transfected defective interfering RNA which contains an IS and a reporter (chloramphenicol acetyltransferase) gene and JHM2c as a helper virus, we demonstrated that subgenomic mRNAs transcribed from the defective interfering RNAs were extremely heterogeneous. The leader-mRNA fusion sites in this virus can be grouped into five types. In type I, the leader is fused with the consensus IS of the template RNA at a site within the UCUAA repeats, consistent with the classical model of discontinuous transcription. In type II, the leader is fused with the consensus IS as in type I, but the leader of mRNA contains some nucleotide substitutions within the UCUAA repeats. In type III, the leader is fused with mRNAs at a site either upstream or downstream of the consensus IS. The sequences around the fusion sites bear little or no homology to the leader. As a result, mRNAs contain sequences complementary to the template sequences upstream of the IS or have sequence deletions downstream of the IS. In type IV, the leader is fused to the IS at the 9-nt sequence immediately downstream of the UCUAA repeats. In type V, the leader-mRNA fusion site contains a duplication of a portion of the leader sequence or an insertion of nontemplated sequences which are not present in either leader or template RNA. These patterns of leader-mRNA fusion resemble the aberrant homologous recombination frequently seen in other RNA viruses. The degree of heterogeneity of leader fusion sites is dependent on the sequences of both the leader RNA and IS. These results suggest that leader-mRNA fusion in coronavirus transcription does not require direct RNA-RNA interaction between complementary sequences. A modified model of RNA transcription and recombination based on protein-RNA and protein-protein interactions is proposed. This study also provides a paradigm for aberrant homologous recombination.

Animals

Hybridization properties of DNA sequences directing the synthesis of messenger RNA and heterogeneous nuclear RNA.

The relationship of the DNA sequences from which polyribosomal messenger RNA (mRNA) and heterogeneous nuclear RNA (NRNA) of mouse L cells are transcribed was investigated by means of hybridization kinetics and thermal denaturation of the hybrids. Hybridization was performed in formamide solutions at DNA excess. Under these conditions most of the hybridizing mRNA and NRNA react at values of D(o)t (DNA concentration multiplied by time) expected for RNA transcribed from the nonrepeated or rarely repeated fraction of the genome. However, a fraction of both mRNA and NRNA hybridize at values of D(o)t about 10,000 times lower, and therefore must be transcribed from highly redundant DNA sequences. The fraction of NRNA hybridizing to highly repeated sequences is about 1.7 times greater than the corresponding fraction of mRNA. The hybrids formed by the rapidly reacting fractions of both NRNA and mRNA melt over a narrow temperature range with a midpoint about 11 degrees C below that of native L cell DNA. This indicates that these hybrids consist of partially complementary sequences with approximately 11% mismatching of bases. Hybrids formed by the slowly reacting fraction of NRNA melt within 4 degrees -6 degrees C of native DNA, indicating very little, if any, mismatching of bases. Hybrids of the slowly reacting components of mRNA, formed under conditions of sufficiently low RNA input, have a high thermal stability, similar to that observed for hybrids of the slowly reacting NRNA component. However, when higher inputs of mRNA are used, hybrids are formed which have a strikingly lower thermal stability. This observation can be explained by assuming that there is sufficient similarity among the relatively rare DNA sequences coding for mRNA so that under hybridization conditions, in which these DNA sequences are not truly in excess, reversible hybrids exhibiting a considerable amount of mispairing are formed. The fact that a comparable phenomenon has not been observed for NRNA may mean that there is less similarity among the relatively rare DNA sequences coding for NRNA than there is among the rare sequences coding for mRNA.

Animals

The Euglena gracilis chloroplast rpoB gene. Novel gene organization and transcription of the RNA polymerase subunit operon.

The rpoB gene coding for a beta-like subunit of the chloroplast DNA-dependent RNA polymerase has been located on the chloroplast genome of Euglena gracilis distal to the rrnC ribosomal RNA operon. We have determined 5760 base-pairs of DNA sequence, including 97 bp of the 5S rRNA gene, an intergenic spacer of 1264 bp, the rpoB gene of 4249 bp, 84 bp spacer and 67 bp of the rpoC1 gene. The rpoB gene is of the same polarity as the rRNA operons. The organization of the rpoB and rpoC genes resembles the E. coli rpoB-rpoC and higher plant chloroplast rpoB-rpoC1-rpoC2 operons. The Euglena rpoB gene (1082 codons) encodes a polypeptide with a predicted molecular weight of 124,288. The rpoB gene is interrupted by seven Group III introns of 93, 95, 94, 99, 101, 110 and 99 bp respectively and a Group II intron of 309 bp. All other known rpoB genes lack introns. All the exon-exon junctions were experimentally determined by cDNA cloning and sequencing or direct primer extension RNA sequencing. Transcripts from the rpoB locus were characterized by Northern hybridization. Fully-spliced, monocistronic rpoB mRNA, as well as rpoB-rpoC1 and rpoB1-rpoC1-rpoC2 mRNAs were identified.

Amino Acid Sequence

Dogme: a nextflow pipeline for reprocessing nanopore RNA and DNA modifications.

MOTIVATION: Oxford Nanopore (ONT) sequencing allows for the direct detection of RNA and DNA modifications from unamplified nucleic acids, which is a significant advantage over other platforms. However, the rapid updates to ONT basecalling models and the evolving landscape of computational tools for modification detection bring about challenges for reproducible and standardized analyses. To address these challenges, we developed Dogme to automate basecalling, alignment, modification detection, and transcript quantification. Dogme automates the reprocessing of ONT POD5 files by integrating basecalling using Dorado, read mapping using minimap2 and subsequent analysis steps such as running modkit. The pipeline supports three major types of sequencing data-direct RNA (dRNA), complementary DNA (cDNA), and genomic DNA (gDNA). Dogme facilitates detection of diverse RNA modifications supported by Dorado such as N6-methyladenosine (m6A), 5-methylcytosine (m5C), inosine, pseudouridine, 2'-O-methylation (Nm) and DNA methylation, while concurrently quantifying full-length transcript isoforms LR-Kallisto for transcript quantification for dRNA and cDNA. RESULTS: We applied Dogme to three separate mouse C2C12 myoblast replicates using direct RNA sequencing on MinION flow cells. We detected 96&#xa0;603 m6A, 43&#xa0;476 m5C, 8829 inosine, 10&#xa0;055 pseudouridine, and 30&#xa0;320 Nm sites in three biological replicates. The pipeline produced reproducible modification profiles and transcript expression levels across replicates, demonstrating its utility for integrative long-read transcriptomic and epigenomic analyses. AVAILABILITY AND IMPLEMENTATION: Dogme is implemented in Nextflow and is freely available under the MIT license at https://github.com/mortazavilab/dogme, with documentation provided for installation and usage.

RNA

Nucleotide sequence at the 3' end of Japanese encephalitis virus genomic RNA.

Japanese encephalitis (JE) virus genomic RNAs were purified from virions. Two hundred nucleotides at the 3' end of JE virus genomic RNA were directly sequenced by using reverse transcriptase. The nucleotide sequence at the 3' end of the viral RNA was conserved among four kinds of JE virus strains. The sequence has no AU-rich region that is present at the 3' termini of alphavirus RNAs. We also compared the nucleotide sequences at the 3' ends of RNAs from three different flaviviruses and found several common sequence elements. A secondary structure at the 3' end of JE virus genomic RNA was proposed that may be common among flavivirus genomic RNAs. Such structures and other common stretches of nucleotide sequences may be related to the biological properties of flaviviruses.

Base Sequence

EIAV genomic organization: further characterization by sequencing of purified glycoproteins and cDNA.

Nucleotide sequence analyses of two different proviral clones of equine infectious anemia virus (EIAV), designated lambda 12 (K. Rushlow et al., 1986, Virology 155, 309-321) and 1369 (T. Kawakami et al., 1987, Virology 158, 300-312), indicate significant differences in the organization of two critical regions of the viral genome, i.e., in the short open reading frames in the pol-env intergenic region and in the 5'-end of the env gene. To determine the correct structure of the EIAV genome, we have performed nucleotide sequence analyses of cDNA clones produced from viral RNA and direct sequencing of purified EIAV envelope glycoproteins (gp90 and gp45). The results of the cDNA sequencing confirm the presence of two short open reading frames in the pol-env intergenic region, as reported previously for the lambda 12 clone. The protein sequencing data correlated exactly with the amino-terminal sequences of gp90 and gp45 deduced from lambda 12 nucleotide sequences. However, the protein sequencing also revealed that the putative signal sequence of EIAV gp90 is not removed during processing. Thus, EIAV apparently contains short open reading frames analogous to human immunodeficiency virus, but differs in its mode of env polyprotein processing.

Amino Acid Sequence

Phased adenine tracts in double-stranded RNA do not induce sequence-directed bending.

Tracts of four to six adenines phased with the DNA helix produce a sequence-directed bending of the helix axis. Here, using gel electrophoresis and electron microscopy (EM), we have asked whether a similar motif will induce bending in a duplex RNA helix. Single-stranded RNAs were transcribed either from short synthetic DNA templates or from Crithidia fasciculata kinetoplast bent DNA, and the complementary single-stranded RNAs were annealed to produce duplex RNA molecules containing blocks of four to six adenines. Electrophoresis on polyacrylamide gels revealed no retardation of the RNAs containing phased blocks of adenines relative to duplex RNAs lacking such blocks. Examination by EM showed most of the molecules to be straight or only slightly bent. Thus, in contrast to DNA duplexes, phased adenine tracts do not induce sequence-directed bending in double-stranded RNA. Analysis of the distribution of molecule shapes for the highly bent C. fasciculata DNA showed that the adenine blocks do not act cooperatively to induce DNA bending and that the molecules must equilibrate between a spectrum of bent shapes.

Animals

Apolipoprotein B mRNA editing is an intranuclear event that occurs posttranscriptionally coincident with splicing and polyadenylation.

The subcellular compartment in which apolipoprotein (apo) B mRNA is edited is unknown. We studied the site of endogenous apoB mRNA editing and correlated the extent of editing with mRNA maturation in the rat liver. RNA editing activity was demonstrated in both nuclear and cytoplasmic extracts. The specific activity of the editing activity was 5.5-fold higher in the nuclear extract, which was not accounted for by activators, inhibitors, or modulators. However, the total editing activity was 3.1 times higher in the cytoplasmic extract. Highly purified rat liver nuclear apoB mRNA contained 17.3 +/- 1.45% edited sequences compared with 56 +/- 2.5% and 62.15 +/- 6.2% edited sequences in hepatic total and polysomal RNAs, respectively. Because of the significant extent of editing of total nuclear RNA, we fractionated it into a poly(A-) and poly(A+) fraction. While the poly(A-) nuclear fraction contained only 10.4 +/- 1.1% edited sequences, which represents a maximum estimate, the poly(A+) nuclear apoB mRNA contained 50 +/- 1.8% edited sequences, a value very similar to that for polysomal RNA. By direct sequencing of cDNA and genomic clones, we found that as in the case of the human apoB gene, the rat apoB gene contains an intron 25 immediately upstream of the edited exon 26. Using this information, we developed a method to examine in a highly selective manner apoB mRNA that is present in the nucleus before splicing of intron 25 and after splicing of this intron. The unspliced nuclear pre-mRNA contained 7.4 +/- 0.2% edited sequences compared with 51.0 +/- 0.9% edited sequences in the spliced nuclear apoB mRNA. Furthermore, in the poly(A-) pool of apoB pre-mRNA, unspliced nuclear pre-mRNA contained hardly any (1.56%) edited sequences, and the spliced nuclear pre-mRNA contained 7.8 +/- 0.6% edited mRNA. In the poly(A+) fraction, unspliced nuclear pre-mRNA had 25.4 +/- 0.05% and spliced nuclear mRNA 53 +/- 0.6% of its apoB mRNA in an edited form. We conclude that in the rat liver apoB mRNA editing is not a cotransciptional event. It occurs posttranscriptionally, but the process is essentially complete in the spliced polyadenylated apoB mRNA before it leaves the nucleus. Little, if any, additional editing occurs in the cytoplasmic compartment.

Animals

Toward the therapeutic editing of mutated RNA sequences.

If RNA editing could be rationally directed to mutated RNA sequences, genetic diseases caused by certain base substitutions could be treated. Here we use a synthetic complementary RNA oligonucleotide to direct the correction of a premature stop codon mutation in dystrophin RNA. The complementary RNA oligonucleotide was hybridized to a premature stop codon and the hybrid was treated with nuclear extracts containing the cellular enzyme double-stranded RNA adenosine deaminase. When the treated RNAs were translated in vitro, a dramatic increase in expression of a downstream luciferase coding region was observed. The cDNA sequence data are consistent with deamination of the adenosine in the UAG stop codon to inosine by double-stranded RNA adenosine deaminase. Injection of oligonucleotide-mRNA hybrids into Xenopus embryos also resulted in an increase in luciferase expression. These experiments demonstrate the principle of therapeutic RNA editing.

Adenosine Deaminase