PubMed Health⌕ Search

Biomedical subjects

Peter N Robinson

Publications and source records attributed to Peter N Robinson.

At least 19 recordsLinked to original sources

Proteomics identify disease-associated variants in patients with rare diseases undiagnosed after genome sequencing.

Despite the introduction of genome sequencing (GS) for rare disease diagnostics, a genetic cause is not identified in most patients. Here, we explored the potential of proteomics to improve the diagnostic yield in 424 patients with rare diseases from the 100,000 Genomes Project (100kGP) without a genetic diagnosis. Serum proteomic profiling was performed using the Olink Explore 1536 assay (N&#xa0;=&#xa0;1463 proteins). For 13 patients without genetic diagnoses, detection of lower serum protein "outliers" (z-score&#xa0;<&#xa0;-2) led to confirmed genetic diagnoses by resolving variants of uncertain significance or prioritizing genes for targeted GS reanalysis. For 23 additional patients without genetic diagnoses (64% of findings), we identified candidate gene-disease links and variants through convergent evidence from lower protein outliers and variants ranked through the variant prioritization tool Exomiser. For example, we identified a candidate heterozygous missense variant [Genome Aggregation Database (gnomAD) minor allele frequency&#xa0;=&#xa0;0.006%] in tyrosine kinase with immunoglobulin-like and epidermal growth factor homology domains 1 (TIE1) that was only present in a patient with lower TIE1 serum abundance (z-score&#xa0;=&#xa0;-5.12) and their father, both of whom were affected by the same monogenic cardiac disorder, but in no other individuals from the 100kGP. Missense (52.5%) and splice region (27.5%) variants accounted for most diagnostic or candidate variants prioritized. This proof-of-principle study demonstrated that serum proteomics can support rare disease diagnosis and identify disease-causing genes in patients undiagnosed after GS, although successful implementation will likely depend on tissue specificity of protein expression, detectability in blood, proteomic platform coverage, and sensitivity.

Humans↗

A phenotypic paradigm for cerebral palsy genetics.

Cerebral palsy (CP) represents a clinically and etiologically heterogeneous group of permanent but not unchanging disorders of movement, posture, and motor function resulting from non-progressive disturbances of the developing fetal or infant brain. Pathogenic variants in Mendelian disease-associated genes can be found in a subset of individuals with CP, with variants deemed causal of CP having been published for at least 515 genes. Currently, controversy exists as to whether to interpret such pathogenic variants as causing CP, whether the diagnosis instead should be "CP mimic," or whether a clinical diagnosis of CP should coexist with the molecular diagnosis of a Mendelian disease. Accordingly, there is no universally accepted model of the genetic architecture of CP. Here, we present a statistical approach that treats CP as a phenotypic feature for which some genetic disorders confer an increased risk. Based on comprehensive literature curation, we show that the null hypothesis of no CP association can be rejected for only 89 of the 515 genes. We applied these findings to the analysis of a cohort of 460 children diagnosed with CP in the Shriner Children's network who underwent genome sequencing. We identified pathogenic or likely pathogenic (P/LP) variants in 60 genes in 15.8% of the children. Only 16 of the 60 genes had significant evidence for CP association in our literature analysis. Our results suggest that a stratified approach to attributing causality to genetic variants in CP could support precision genomic medicine for affected individuals.

Humans↗

Oncopacket: integration of cancer research data using GA4GH phenopackets.

SUMMARY: Lack of data integration remains a significant impediment to cancer research, and many analyses still require customized software to transform and prepare cancer data. We describe a software package to harmonize genetic and clinical cancer data into the GA4GH Phenopacket schema, an ISO standard for representing clinical case data. We integrated demographic, mutation, morphology, diagnosis, intervention, and survival data using case data from the National Cancer Institute for 12 cancer types. The Phenopacket standard provides a foundation for downstream use, including sophisticated statistical and AI/ML analyses. We demonstrate fitness for purpose by using the integrated data to recapitulate a known association between mutations in the gene encoding isocitrate dehydrogenase 1 and survival time in brain cancer patients. AVAILABILITY AND IMPLEMENTATION: Source code is freely available at: https://github.com/monarch-initiative/oncopacket (archived at 10.5281/zenodo.15353125).

Humans↗

Leveraging clinical intuition to improve accuracy of phenotype-driven prioritization.

PURPOSE: Clinical intuition is commonly incorporated into the differential diagnosis as an assessment of the likelihood of candidate diagnoses based either on the patient population being seen in a specific clinic or on the signs and symptoms of the initial presentation. Algorithms to support diagnostic sequencing in individuals with a suspected rare genetic disease do not yet incorporate intuition and instead assume that each Mendelian disease has an equal pretest probability. METHODS: The LIkelihood Ratio Interpretation of Clinical AbnormaLities (LIRICAL) algorithm calculates the likelihood ratio of clinical manifestations represented by Human Phenotype Ontology terms to rank candidate diagnoses. The initial version of LIRICAL assumed an equal pretest probability for each disease in its calculation of the posttest probability (where the test is diagnostic exome or genome sequencing). We introduce Clinical Intuition for Likelihood Ratios (ClintLR), an extension of the LIRICAL algorithm that boosts the pretest probability of groups of related diseases deemed to be more likely. RESULTS: The average rank of the correct diagnosis in simulations using ClintLR showed a statistically significant improvement over a range of adjustment factors. CONCLUSION: ClintLR successfully encodes clinical intuition to improve ranking of rare diseases in diagnostic sequencing. ClintLR is freely available at https://github.com/TheJacksonLaboratory/ClintLR.

Humans↗

A corpus of GA4GH phenopackets: Case-level phenotyping for genomic diagnostics and discovery.

The Global Alliance for Genomics and Health (GA4GH) Phenopacket Schema was released in 2022 and approved by ISO as a standard for sharing clinical and genomic information about an individual, including phenotypic descriptions, numerical measurements, genetic information, diagnoses, and treatments. A phenopacket can be used as an input file for software that supports phenotype-driven genomic diagnostics and for algorithms that facilitate patient classification and stratification for identifying new diseases and treatments. There has been a great need for a collection of phenopackets to test software pipelines and algorithms. Here, we present Phenopacket Store. Phenopacket Store v.0.1.19 includes 6,668 phenopackets representing 475 Mendelian and chromosomal diseases associated with 423 genes and 3,834 unique pathogenic alleles curated from 959 different publications. This represents the first large-scale collection of case-level, standardized phenotypic information derived from case reports in the literature with detailed descriptions of the clinical data and will be useful for many purposes, including the development and testing of software for prioritizing genes and diseases in diagnostic genomics, machine learning analysis of clinical phenotype data, patient stratification, and genotype-phenotype correlations. This corpus also provides best-practice examples for curating literature-derived data using the GA4GH Phenopacket Schema.

Humans↗

Induction of macrophage chemotaxis by aortic extracts of the mgR Marfan mouse model and a GxxPG-containing fibrillin-1 fragment.

BACKGROUND: The primary cause of early death in untreated Marfan syndrome (MFS) patients is aortic dilatation and dissection. METHODS AND RESULTS: We investigated whether ascending aortic samples from the fibrillin-1-underexpressing mgR mouse model for MFS or a recombinant fibrillin-1 fragment containing an elastin-binding protein (EBP) recognition sequence can act as chemotactic stimuli for macrophages. Both the aortic extracts from the mgR/mgR mice and the fibrillin-1 fragment significantly increased macrophage chemotaxis compared with extracts from wild-type mice or buffer controls. The chemotactic response was significantly diminished by pretreatment of macrophages with lactose or with the elastin-derived peptide VGVAPG and by pretreatment of samples with a monoclonal antibody directed against an EBP recognition sequence. Mutation of the EBP recognition sequence in the fibrillin-1 fragment also abolished the chemotactic response. These results indicate the involvement of EBP in mediating the effects. Additionally, investigation of macrophages in aortic specimens of MFS patients demonstrated macrophage infiltration in the tunica media. CONCLUSIONS: Our findings demonstrate that aortic extracts from mgR/mgR mice can stimulate macrophage chemotaxis by interaction with EBP and show that a fibrillin-1 fragment possesses chemotactic stimulatory activity similar to that of elastin degradation peptides. They provide a plausible molecular mechanism for the inflammatory infiltrates observed in the mgR mouse model and suggest that inflammation may represent a component of the complex pathogenesis of MFS.

Animals↗

Gene identification and analysis of transcripts differentially regulated in fracture healing by EST sequencing in the domestic sheep.

BACKGROUND: The sheep is an important model animal for testing novel fracture treatments and other medical applications. Despite these medical uses and the well known economic and cultural importance of the sheep, relatively little research has been performed into sheep genetics, and DNA sequences are available for only a small number of sheep genes. RESULTS: In this work we have sequenced over 47 thousand expressed sequence tags (ESTs) from libraries developed from healing bone in a sheep model of fracture healing. These ESTs were clustered with the previously available 10 thousand sheep ESTs to a total of 19087 contigs with an average length of 603 nucleotides. We used the newly identified sequences to develop RT-PCR assays for 78 sheep genes and measured differential expression during the course of fracture healing between days 7 and 42 postfracture. All genes showed significant shifts at one or more time points. 23 of the genes were differentially expressed between postfracture days 7 and 10, which could reflect an important role for these genes for the initiation of osteogenesis. CONCLUSION: The sequences we have identified in this work are a valuable resource for future studies on musculoskeletal healing and regeneration using sheep and represent an important head-start for genomic sequencing projects for Ovis aries, with partial or complete sequences being made available for over 5,800 previously unsequenced sheep genes.

Animals↗

Escobar syndrome is a prenatal myasthenia caused by disruption of the acetylcholine receptor fetal gamma subunit.

Escobar syndrome is a form of arthrogryposis multiplex congenita and features joint contractures, pterygia, and respiratory distress. Similar findings occur in newborns exposed to nicotinergic acetylcholine receptor (AChR) antibodies from myasthenic mothers. We performed linkage studies in families with Escobar syndrome and identified eight mutations within the gamma -subunit gene (CHRNG) of the AChR. Our functional studies show that gamma -subunit mutations prevent the correct localization of the fetal AChR in human embryonic kidney-cell membranes and that the expression pattern in prenatal mice corresponds to the human clinical phenotype. AChRs have five subunits. Two alpha, one beta, and one delta subunit are always present. By switching gamma to epsilon subunits in late fetal development, fetal AChRs are gradually replaced by adult AChRs. Fetal and adult AChRs are essential for neuromuscular signal transduction. In addition, the fetal AChRs seem to be the guide for the primary encounter of axon and muscle. Because of this important function in organogenesis, human mutations in the gamma subunit were thought to be lethal, as they are in gamma -knockout mice. In contrast, many mutations in other subunits have been found to be viable but cause postnatally persisting or beginning myasthenic syndromes. We conclude that Escobar syndrome is an inherited fetal myasthenic disease that also affects neuromuscular organogenesis. Because gamma expression is restricted to early development, patients have no myasthenic symptoms later in life. This is the major difference from mutations in the other AChR subunits and the striking parallel to the symptoms found in neonates with arthrogryposis when maternal AChR auto-antibodies crossed the placenta and caused the transient inactivation of the AChR pathway.

Amino Acid Sequence↗

A novel mutation in the DNA-binding domain of MAF at 16q23.1 associated with autosomal dominant "cerulean cataract" in an Indian family.

Congenital cataract, a clinically and genetically highly heterogeneous eye disorder, is one of the significant causes of visual impairment or blindness in children. It is frequently inherited as an autosomal dominant trait. We investigated a three-generation family of Indian origin with 12 members affected with cerulean cataract. Linkage analysis was carried out in this family using more than 100 microsatellite markers for the known cataract candidate gene loci. A positive two-point lod score of 3.9 at theta = 0.000, indicative of linkage, was obtained with three microsatellite markers for chromosome 16. Multipoint and haplotype analysis narrowed the cataract locus to a 15.3 cM region between markers D16S518 and D16S511 that corresponds to the region 16q23.1. Direct sequencing of the candidate gene MAF, which lies in the critical linked region, revealed a novel heterozygous missense mutation in the basic region (BR) of the DNA-binding domain. This sequence change was considered pathogenic as it segregated in all affected family members, neither seen in unaffected family members nor in 106 unrelated controls. The mutation also results in substitution of highly conserved lysine 297 by arginine (K297R) that affects a residue that forms a part of a predicted DNA-interaction region of the protein. The association of microcornea with congenital cataract in some affected individuals further underlines the role of the MAF transcription factor in lens and anterior ocular development. Our findings expand the mutation spectrum of MAF in association with congenital cataract and highlight the genetic and phenotypic heterogeneity of congenital cataract.

Adolescent↗

HotSwap for bioinformatics: a STRAP tutorial.

BACKGROUND: Bioinformatics applications are now routinely used to analyze large amounts of data. Application development often requires many cycles of optimization, compiling, and testing. Repeatedly loading large datasets can significantly slow down the development process. We have incorporated HotSwap functionality into the protein workbench STRAP, allowing developers to create plugins using the Java HotSwap technique. RESULTS: Users can load multiple protein sequences or structures into the main STRAP user interface, and simultaneously develop plugins using an editor of their choice such as Emacs. Saving changes to the Java file causes STRAP to recompile the plugin and automatically update its user interface without requiring recompilation of STRAP or reloading of protein data. This article presents a tutorial on how to develop HotSwap plugins. STRAP is available at http://strapjava.de and http://www.charite.de/bioinf/strap. CONCLUSION: HotSwap is a useful and time-saving technique for bioinformatics developers. HotSwap can be used to efficiently develop bioinformatics applications that require loading large amounts of data into memory.

Algorithms↗

Novel sequence variants in dysferlin-deficient muscular dystrophy leading to mRNA decay and possible C2-domain misfolding.

Mutations in the gene encoding dysferlin (DYSF) cause the allelic autosomal recessive disorders limb girdle muscular dystrophy 2B and Miyoshi myopathy. It encompasses 55 exons spanning 150 kb of genomic DNA. Dysferlin is involved in membrane repair in skeletal muscle. We identified three families with novel sequence variants in DYSF. All affected family members showed limb girdle weakness and had reduced or absent dysferlin protein on immunohistochemistry. All exons of DYSF were screened by genomic sequencing. Five novel variants in DYSF were found: two missense mutations (c.895G>A and c.4022T>C), one 5' donor splice-site variant (c.855+1delG), one nonsense mutation (c.1448C>A), and a variant in the 3'UTR of DYSF (c.*107T>A). All alterations were confirmed by restriction enzyme analysis and not found in 400 control alleles. Nonsense mediated RNA decay or changes in the three-dimensional protein structure resulting in intracellular dysferlin aggregates and finally the lack of dysferlin protein were identified as consequences of the novel DYSF variants.

Adult↗

A fibrillin-1-fragment containing the elastin-binding-protein GxxPG consensus sequence upregulates matrix metalloproteinase-1: biochemical and computational analysis.

Mutations in the gene for fibrillin-1 cause Marfan syndrome (MFS), a common hereditary disorder of connective tissue. Recent findings suggest that proteolysis, increased matrix metalloproteinase activity, and fragmentation of fibrillin-rich microfibrils in tissues of persons with MFS contribute to the complex pathogenesis of this disorder. In this study we show that a fibrillin-1 fragment containing a EGFEPG sequence that conforms to a putative GxxPG elastin-binding protein (EBP) consensus sequence upregulates the expression and production of matrix metalloproteinase (MMP)-1 by up to ninefold in a cell culture system. A mutation of the GxxPG consensus sequence site abrogated the effects. This is the first demonstration of such an effect for ligands other than elastin fragments. Molecular dynamics analysis of oligopeptides with the wildtype and mutant sequence support our biochemical results by predicting significant alterations of structural characteristics such as the potential for forming a type VIII beta-turn that are thought to be important for binding to the EBP. These results suggest that fibrillin-1 fragments may regulate MMP-1 expression, and that the dysregulation of MMPs related to fragmentation of fibrillin might contribute to the development of MFS. Our Gene Ontology (GO) analysis of the human proteome shows that proteins with multiple GxxPG motifs are highly enriched for GO terms related to the extracellular matrix. Matrix proteins with multiple GxxPG sites include fibrillin-1, -2, and -3, elastin, fibronectin, laminin, and several tenascins and collagens. Some of these proteins have been associated with disorders involving alterations in MMP regulation, and the results of the present study suggest a potential mechanism for these observations.

Amino Acid Sequence↗

Binary state pattern clustering: a digital paradigm for class and biomarker discovery in gene microarray studies of cancer.

Class and biomarker discovery continue to be among the preeminent goals in gene microarray studies of cancer. We have developed a new data mining technique, which we call Binary State Pattern Clustering (BSPC) that is specifically adapted for these purposes, with cancer and other categorical datasets. BSPC is capable of uncovering statistically significant sample subclasses and associated marker genes in a completely unsupervised manner. This is accomplished through the application of a digital paradigm, where the expression level of each potential marker gene is treated as being representative of its discrete functional state. Multiple genes that divide samples into states along the same boundaries form a kind of gene-cluster that has an associated sample-cluster. BSPC is an extremely fast deterministic algorithm that scales well to large datasets. Here we describe results of its application to three publicly available oligonucleotide microarray datasets. Using an alpha-level of 0.05, clusters reproducing many of the known sample classifications were identified along with associated biomarkers. In addition, a number of simulations were conducted using shuffled versions of each of the original datasets, noise-added datasets, as well as completely artificial datasets. The robustness of BSPC was compared to that of three other publicly available clustering methods: ISIS, CTWC and SAMBA. The simulations demonstrate BSPC's substantially greater noise tolerance and confirm the accuracy of our calculations of statistical significance.

Algorithms↗

Calcium-dependent self-association of the C-type lectin domain of versican.

Versican is a large (1-2 x 10(6) Da) chondroitin-sulfate proteoglycan that can form large aggregates by means of interaction with hyaluronan and also binds to a series of other extracellular matrix proteins, chemokines and cell-surface molecules. Versican is a multifunctional molecule with roles in cell adhesion, matrix assembly, cell migration and proliferation. Characterization of the binding interactions mediated by the various domains of versican is a first step towards understanding the functions of versican and interacting molecules in the extracellular matrix. In this study we investigated a recombinant construct corresponding to the C-type lectin domain of versican and demonstrated a calcium-dependent self-association of this region by blot overlay and plasmon surface resonance assays. Electron microscopy provided further evidence of the relevance of the binding reaction by demonstrating a mixture of monomers, dimers and complex aggregates of recombinant versican C-type lectin domain. This binding reaction could contribute to the ability of versican to organize formation of the proteoglycan extracellular matrix by inducing binding of individual versican molecules or by modulating binding reactions to other matrix components.

Calcium↗

Shprintzen-Goldberg syndrome: fourteen new patients and a clinical analysis.

The Shprintzen-Goldberg syndrome (SGS) is a disorder of unknown cause comprising craniosynostosis, a marfanoid habitus and skeletal, neurological, cardiovascular, and connective-tissue anomalies. There are no pathognomonic signs of SGS and diagnosis depends on recognition of a characteristic combination of anomalies. Here, we describe 14 persons with SGS and compare their clinical findings with those of 23 previously reported individuals, including two families with more than one affected individual. Our analysis suggests that there is a characteristic facial appearance, with more than two thirds of all individuals having hypertelorism, down-slanting palpebral fissures, a high-arched palate, micrognathia, and apparently low-set and posteriorly rotated ears. Other commonly reported manifestations include hypotonia in at least the neonatal period, developmental delay, and inguinal or umbilical hernia. The degree of reported intellectual impairment ranges from mild to severe. The most common skeletal manifestations in SGS were arachnodactyly, pectus deformity, camptodactyly, scoliosis, and joint hypermobility. None of the skeletal signs alone is specific for SGS. Our study includes 14 mainly German individuals with SGS evaluated over a period of 10 years. Given that only 23 other persons with SGS have been reported to date worldwide, we suggest that SGS may be more common than previously assumed.

Abnormalities, Multiple↗

RGD-containing fibrillin-1 fragments upregulate matrix metalloproteinase expression in cell culture: a potential factor in the pathogenesis of the Marfan syndrome.

The Marfan syndrome (MFS), a relatively common autosomal dominant disorder of connective tissue, is caused by mutations in the gene for fibrillin-1 (FBN1). Fibrillin-1 is the main component of the 10- to 12-nm microfibrils that together with elastin form elastic fibers found in tissues such as the aortic media. Recently, FBN1 mutations have been shown to increase the susceptibility of fibrillin-1 to proteolysis in vitro, and other findings suggest that up-regulation of matrix metalloproteinases (MMP), as well as fragmentation of microfibrils, could play a role in the pathogenesis of MFS. In the present work, we have investigated the influence of fibrillin-1 fragments on the expression of MMP-1, MMP-2, and MMP-3 in a cell culture system. Cultured human dermal fibroblasts were incubated with several different recombinant fibrillin-1 fragments. The expression level of MMP-1, MMP-2, and MMP-3, was determined by quantitative reverse transcriptase-polymerase chain reaction (RT-PCR), and the concentration of the corresponding proteins was estimated by quantitative Western blotting. Our results establish that treatment of cultured human dermal fibroblasts with recombinant fibrillin-1 fragments containing the arginine-glycine-aspartic acid (RGD) integrin-binding motif of fibrillin-1 induces up-regulation of MMP-1 and MMP-3. A similar effect was seen upon stimulation with a synthetic RGD peptide. The expression of MMP-2 was not influenced by treatment. Our results suggest the possibility that fibrillin fragments could themselves have pathogenic effects by leading to up-regulation of MMPs, which in turn may be involved in the progressive breakdown of microfibrils thought to play a role in MFS.

Enzyme Induction↗

A molecular pathogenesis for transcription factor associated poly-alanine tract expansions.

Poly-alanine (Ala) tract expansions in transcription factors have been shown to be associated with human birth defects such as malformations of the brain, the digits, and other structures. Expansions of a poly-Ala tract from 15 to 22 (+7)-29 (+14) Ala in Hoxd13, for example, result in the limb malformation synpolydactyly in humans and in mice [synpolydactyly homolog (spdh)]. Here, we show that an increase of the Ala repeat above a certain length (22 Ala) is associated with a shift in the localization of Hoxd13 from nuclear to cytoplasmic, where it forms large amorphous aggregates. We observed similar aggregates for expansion mutations in SOX3, RUNX2 and HOXA13, pointing to a common mechanism. Cytoplasmic aggregation of mutant Hoxd13 protein is influenced by the length of the repeat, the level of expression and the efficacy of degradation by the proteasome. Heat shock proteins Hsp70 and Hsp40 co-localize with the aggregates and activation of the chaperone system by geldanamycin leads to a reduction of aggregate formation. Furthermore, recombinant mutant Hoxd13 protein forms aggregates in vitro demonstrating spontaneous misfolding of the protein. We analyzed the mouse mutant spdh, which harbors a +7 Ala expansion in Hoxd13 similar to the human synpolydactyly mutations, as an in vivo model and were able to show a reduction of mutant Hoxd13 and, in contrast to wt Hoxd13, a primarily cytoplasmic localization of the protein. Our results provide evidence that poly-Ala repeat expansions in transcription factors result in misfolding, degradation and cytoplasmic aggregation of the mutant proteins.

Animals↗

Gene-Ontology analysis reveals association of tissue-specific 5' CpG-island genes with development and embryogenesis.

A key open question in the understanding of the biology of DNA methylation relates to the origin and function of CpG islands, stretches of GC-rich and relatively CpG-rich DNA sequence that often colocalize with promoter regions. All housekeeping, but also a substantial minority of tissue-specific genes are associated with the CpG islands. Limited experimental evidence suggests that CpG islands are associated with promoters or replication origins active during early development. Although this hypothesis is attractive for widely expressed genes, which would be expected to be expressed during early development, many tissue-specific genes also contain CpG islands. In this work, we used a genome-wide Gene-Ontology (GO)-based approach to analyze associations between GO terms and the presence of 5' CpG islands in human Reference Sequence (RefSeq) genes. We found that 19 of the 3849 GO terms with at least one annotated human sequence showed a highly significant association with the likelihood of 5' CpG islands being present in genes annotated to that term. In particular, the term 'development' showed a highly significantly increased proportion of 5' CpG island genes. The overrepresentation of 5' CpG island genes was even more significant for tissue-specific RefSeqs annotated to development as well as many of its descendent terms. In addition, the proportion of expressed sequence tags from embryonic libraries amongst tissue-specific genes was twice as high for RefSeqs with 5' CpG islands as for those without CpG islands. These results provide strong support for previous speculations that early embryonic expression is associated with CpG islands.

Animals↗