PubMed Health⌕ Search

Biomedical subjects

Lluis Armengol

Publications and source records attributed to Lluis Armengol.

6 recordsLinked to original sources

Global variation in copy number in the human genome.

Copy number variation (CNV) of DNA sequences is functionally significant but has yet to be fully ascertained. We have constructed a first-generation CNV map of the human genome through the study of 270 individuals from four populations with ancestry in Europe, Africa or Asia (the HapMap collection). DNA from these individuals was screened for CNV using two complementary technologies: single-nucleotide polymorphism (SNP) genotyping arrays, and clone-based comparative genomic hybridization. A total of 1,447 copy number variable regions (CNVRs), which can encompass overlapping or adjacent gains or losses, covering 360 megabases (12% of the genome) were identified in these populations. These CNVRs contained hundreds of genes, disease loci, functional elements and segmental duplications. Notably, the CNVRs encompassed more nucleotide content per genome than SNPs, underscoring the importance of CNV in genetic diversity and evolution. The data obtained delineate linkage disequilibrium patterns for many CNVs, and reveal marked variation in copy number among populations. We also demonstrate the utility of this resource for genetic disease studies.

Chromosome Mapping↗

Genome assembly comparison identifies structural variants in the human genome.

Numerous types of DNA variation exist, ranging from SNPs to larger structural alterations such as copy number variants (CNVs) and inversions. Alignment of DNA sequence from different sources has been used to identify SNPs and intermediate-sized variants (ISVs). However, only a small proportion of total heterogeneity is characterized, and little is known of the characteristics of most smaller-sized (<50 kb) variants. Here we show that genome assembly comparison is a robust approach for identification of all classes of genetic variation. Through comparison of two human assemblies (Celera's R27c compilation and the Build 35 reference sequence), we identified megabases of sequence (in the form of 13,534 putative non-SNP events) that were absent, inverted or polymorphic in one assembly. Database comparison and laboratory experimentation further demonstrated overlap or validation for 240 variable regions and confirmed >1.5 million SNPs. Some differences were simple insertions and deletions, but in regions containing CNVs, segmental duplication and repetitive DNA, they were more complex. Our results uncover substantial undescribed variation in humans, highlighting the need for comprehensive annotation strategies to fully interpret genome scanning and personalized sequencing projects.

Base Sequence↗

Complex patterns of copy number variation at sites of segmental duplications: an important category of structural variation in the human genome.

The structural diversity of the human genome is much higher than previously assumed although its full extent remains unknown. To investigate the association between segmental duplications that display constitutive copy number differences (CNDs) between humans and the great apes and those which exhibit polymorphic copy number variations (CNVs) between humans, we analysed a BAC array enriched with segmental duplications displaying such CNDs. This study documents for the first time that in addition to human-specific gains common to all humans, these duplication clusters (DCs) also exhibit polymorphic CNVs > 40 kb. Segmental duplication is known to have been a frequent event during human genome evolution. Importantly, among the CNV-associated genes identified here, those involved in transcriptional regulation were found to be significantly overrepresented. Complex patterns of variation were evident at sites of DCs, manifesting as inter-individual differentially sized copy number alterations at the same genomic loci. Thus, CNVs associated with segmental duplications do not simply represent insertion/deletion polymorphisms, but rather constitute a wide variety of rearrangements involving differential amplification and partial gains and losses with high inter-individual variability. Although the number of CNVs was not found to differ between Africans and Caucasians/Asians, the average number of variant patterns per locus was significantly lower in Africans. Thus, complex variation patterns characterizing segmental duplications result from relatively recent genomic rearrangements. The high number of these rearrangements, some of which are potentially recurrent, together with differences in population size and expansion dynamics, may account for the greater diversity of CNV in Caucasians/Asians as compared with Africans.

Animals↗

Identification of large-scale human-specific copy number differences by inter-species array comparative genomic hybridization.

Copy number differences (CNDs), and the concomitant differences in gene number, have contributed significantly to the genomic divergence between humans and other primates. To assess its relative importance, the genomes of human, common chimpanzee, bonobo, gorilla, orangutan and macaque were compared by comparative genomic hybridization using a high-resolution human BAC array (aCGH). In an attempt to avoid potential interference from frequent intra-species polymorphism, pooled DNA samples were used from each species. A total of 322 sites of large-scale inter-species CND were identified. Most CNDs were lineage-specific but frequencies differed considerably between the lineages; the highest CND frequency among hominoids was observed in gorilla. The conserved nature of the orangutan genome has already been noted by karyotypic studies and our findings suggest that this degree of conservation may extend to the sub-microscopic level. Of the 322 CND sites identified, 14 human lineage-specific gains were observed. Most of these human-specific copy number gains span regions previously identified as segmental duplications (SDs) and our study demonstrates that SDs are major sites of CND between the genomes of humans and other primates. Four of the human-specific CNDs detected by aCGH map close to the breakpoints of human-specific karyotypic changes [e.g., the human-specific inversion of chromosome 1 and the polymorphic inversion inv(2)(p11.2q13)], suggesting that human-specific duplications may have predisposed to chromosomal rearrangement. The association of human-specific copy number gains with chromosomal breakpoints emphasizes their potential importance in mediating karyotypic evolution as well as in promoting human genomic diversity.

Animals↗

Enrichment of segmental duplications in regions of breaks of synteny between the human and mouse genomes suggest their involvement in evolutionary rearrangements.

The sequence of the mouse genome allows one to compare the conservation of synteny between the human and mouse genome and exploration of regions that might have been involved in major rearrangements during the evolution of these two species (evolutionary genome rearrangements). Recent segmental duplications (or duplicons) are paralogous DNA sequences with high sequence identity that account for about 3.5-5% of the human genome and have emerged during the past approximately 35 million years of evolution. These regions are susceptible to illegitimate recombination leading to rearrangements that result in genomic disorders or genomic mutations. A catalogue of several hundred segmental duplications potentially leading to genomic rearrangements has been reported. The authors and others have observed that some chromosome regions involved in genomic disorders are shuffled in orientation and order in the mouse genome and that regions flanked by segmental duplications are often polymorphic. We have compared the human and mouse genome sequences and demonstrate here that recent segmental duplications correlate with breaks of synteny between these two species. We also observed that nine primary regions involved in human genomic disorders show changes in the order or the orientation of mouse/human synteny segments, were often flanked by segmental duplications in the human sequence. We found that 53% of all evolutionary rearrangement breakpoints associate with segmental duplications, as compared with 18% expected in a random location of breaks along the chromosome (P<0.0001). Our data suggest that segmental duplications have participated in the recent evolution of the human genome, as driving forces for evolutionary rearrangements, chromosome structure polymorphisms and genomic disorders.

Animals↗

Human chromosome 7: DNA sequence and biology.

DNA sequence and annotation of the entire human chromosome 7, encompassing nearly 158 million nucleotides of DNA and 1917 gene structures, are presented. To generate a higher order description, additional structural features such as imprinted genes, fragile sites, and segmental duplications were integrated at the level of the DNA sequence with medical genetic data, including 440 chromosome rearrangement breakpoints associated with disease. This approach enabled the discovery of candidate genes for developmental diseases including autism.

Animals↗