PubMed Health⌕ Search

PubMed · 15777644

Identification and characterization of coding single-nucleotide polymorphisms within human protocadherin-alpha and -beta gene clusters.

Abstract

The human protocadherin (Pcdh) gene clusters are located on chromosome 5q31. Single-nucleotide polymorphisms (SNPs) were detected in the Pcdh-alpha and -beta variable exons, and in the Pcdh-alpha constant exon, in samples from 104 individuals. Among coding SNPs (cSNPs), nonsynonymous (amino acid exchange) SNPs were 2.2 times more common than synonymous (silent) changes in the Pcdh-alpha variable exons, but only 1.2 times more common in the Pcdh-beta variable exons. The nonsynonymous SNPs were high in the ectodomain (EC) 1 encoding region of Pcdh-alpha but not of Pcdh-beta. One 48-kb region of extensive linkage disequilibrium (LD) is reported that has two haplotypes extending from the alpha1 to alpha7 genes in the Pcdh-alpha cluster. Here we identified 15 amino acid exchanges in these two major haplotypes; therefore, the two haplotypes encode different sets of Pcdh-alpha proteins in the brain. The distribution of cSNPs was different for each EC region of Pcdh-alpha or -beta. The frequency of cSNPs was negatively correlated with the paralogous sequence diversity. These results suggested that gene conversion events in homologous regions of the Pcdh-alpha and Pcdh-beta clusters generated the cSNPs. Within the cSNPs, gene conversions were found in Pcdh-alpha4 in the major haplotype, and in Pcdh-beta9. These gene conversions were caused by the unequal crossing-over of homologous sequence regions. Thus, nonsynonymous variations in the Pcdh-alpha and -beta genes are possible contributors to the variations in human brain function.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Rie Miki, Kotaro Hattori, Yusuke Taguchi, Motoki N Tada, Tomoko Isosaka, Yuko Hidaka, Takahiro Hirabayashi, Ryota Hashimoto, Hiroshi Fukuzako, Takeshi Yagi. 2005-04-11. Identification and characterization of coding single-nucleotide polymorphisms within human protocadherin-alpha and -beta gene clusters.. https://doi.org/10.1016/j.gene.2004.11.044

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

A de novo algorithm for allele reconstruction from Oxford nanopore amplicon reads, with application to CYP2D6.

MOTIVATION: The Oxford Nanopore Technologies' sequencing platform offers a path towards bedside genomics, producing long reads that can completely cover a gene of interest, and detect any known or novel variant the gene contains. However, the analysis of these long reads to identify actionable genotypes remains challenging and typically requires customization depending on the target gene. RESULTS: Here, we describe a generic algorithm to accurately reconstruct allele sequences derived from long-reads of amplicon-based data. Rather than calling variants directly from these long-reads, our method takes a "sequence-first" approach, performing an unbiased reconstruction of the underlying amplicon sequences to generate high-confidence reconstructed allele sequences. This is done without user input of the target gene, allowing for any source amplicon to be reconstructed. These high-confidence reconstructed allele sequences are then compared to the genomic reference sequence of the gene to infer the specific diplotype present in the sample. This approach is agnostic towards the number of genes and alleles present and readily detects novel variants. We demonstrate our approach using three independent data sets for CYP2D6, a diverse and complex gene with over 175 known alleles of clinical significance. We show how our approach can accurately recover validated CYP2D6 diplotypes from 20 Coriell samples covering 14 distinct alleles, using different amplicons, flow cell versions, and depths. This includes inferring occurrences of allele duplication events from relative abundances of each allele, a critical factor for ascribing functional effects to a diplotype. Further, we demonstrate our approach's utility for other genomic regions, including HLA. AVAILABILITY: Custom code is available at the following GitHub repository, along with instructions for use and test data: https://github.com/scottdbrown/allele-reconstruction-long-read-amplicon-data. A snapshot of the code at the time of publication is available on Zenodo.org; doi 10.5281/zenodo.19716004. Raw .fastq sequence data for our three sequencing runs is available at the SRA under Bioproject PRJNA1357883 (https://www.ncbi.nlm.nih.gov/bioproject/1357883).

Alleles↗

Supergene control of chiral development in mirror-image flowers.

How genes determine the development of chiral structures is a fascinating question. The reciprocal placement of female and male organs on opposite sides of mirror-image flowers promotes efficient cross-pollination. Here, we identified that in butterfly lilies, female and male organs deflect by a combination of genetically controlled chirality and gravitropism, orienting left and right with respect to an external rather than internal reference axis. We found coordinated organ placement to be controlled by a hemizygous supergene containing two candidate causal loci, MIR156-R and YUCCA-R, that are responsible for opposite female and male organ orientation, respectively. The resulting differential placement of pollen carrying the two supergene alleles on pollinators' bodies leads to their transfer to the stigmas of flowers with opposite handedness and maintenance of the reproductive polymorphism.

Alleles↗

DirectASRM: uncovering allele-specific post-transcriptional RNA modifications through direct RNA sequencing.

SUMMARY: We developed DirectASRM, a comprehensive database for the systematic identification, integration, and annotation of allele-specific RNA modifications (ASRMs) from direct RNA sequencing data. DirectASRM enables single-base, transcript-level detection of ASRMs across multiple RNA modification types, diverse organisms and condition-specific contexts. The database further evaluates the confidence of each ASRM-SNP pair association within isoform context by jointly considering statistical evidence of allelic modification imbalance and independent support from external next-generation sequencing (NGS) - based RNA modification resources. DirectASRM also provides extensive functional annotations for ASRMs and their associated variants, including intra-sample transcript-level allele-specific expression (ASE) and allele-specific splicing, as well as additional post-transcriptional regulatory features such as miRNA binding, circRNA, RNA-protein interactions, and disease relevance. Overall, DirectASRM serves as a comprehensive resource that supports systematic investigation of the potential functional impact of genetic variants in epitranscriptomic regulation. AVAILABILITY AND IMPLEMENTATION: DirectASRM database is freely accessible at http://modinfor.com/DirectASRM/. DirectASRM pipeline is available at GitHub (https://github.com/jiayin1101/DirectASRM_pipeline) and Zenodo (DOI: https://doi.org/10.5281/zenodo.19876077).

Alleles↗