PubMed Health⌕ Search

PubMed · 9806450

The Human Genome Project: from mapping to sequencing.

Abstract

Until recently, the "human genome" programs were mainly directed towards the development of maps to identify disease genes. The genetic map comprises about 8000 highly informative second generation markers of the microsatellite type. The density of markers is now sufficient to localize a gene for a monogenic disease with a precision of 1 to 2 million base pairs easily, and to define intervals which contain susceptibility genes for multifactorial disorders. A third generation map based on single nucleotide polymorphisms that can be genotyped using DNA chip technology is in progress. The physical map, based on sets of overlapping yeast artificial chromosomes ordered using sequence-tagged sites, covers over 90% of the genome. However, this physical map cannot serve as a support for sequencing because of the numerous rearrangements that occur in yeast artificial chromosomes. An international network of laboratories has mapped a set of more than 30,000 expressed sequences from cDNAs using whole genome radiation hybrids that enable integration of genes within existing maps. The human genome program is now progressively shifting to massive sequencing, although sequence ready maps are not available for the major part of the human genome. Similarly, our capacity to interpret the available genomic sequence remains limited.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

J Weissenbach. 1998. The Human Genome Project: from mapping to sequencing.. https://doi.org/10.1515/cclm.1998.086

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Statistical test to compare the linkage model and the admixture model based on central limit results.

In the Admixture Model, the probability that an individual carries a certain allele at a specific marker depends on the allele frequencies in K ancestral populations and the proportion of the individual's genome originating from these populations. The markers are assumed to be independent. The Linkage Model is a Hidden Markov Model that extends the Admixture Model by incorporating linkage between neighboring loci. We prove consistency and asymptotic normality of maximum likelihood estimators for the ancestry of individuals in the Linkage Model, complementing earlier results by (Pfaff et al., 2004; Pfaffelhuber and Rohde, 2022; Heinzel, 2025) for the Admixture Model. These results are used to prove that a statistical test that allows for model selection between the Admixture Model and the Linkage Model is an asymptotic level-α-test. Finally, we demonstrate the practical relevance of our results by applying the test to real-world data from The 1000 Genomes Project Consortium (2015).

Genetic Linkage↗

Haplotype thinking in lung disease.

To identify the genetic etiology of a disease of interest, disease-related characteristics (phenotypes) are often tested for association with genetic variants (genotypes). Although genetic association studies of single genetic variants have been widely performed, there has been increasing interest in studies of multiple adjacent genetic variants on one chromosome, known as a haplotype. In this review, we will provide background about the origin of haplotypes and why they can be useful in genetic studies; we will discuss approaches to determining haplotypes and performing haplotype-based genetic association studies; and we will compare single variant and haplotype-based approaches.

Genetic Linkage↗

Associations between DNA markers and resistance to diseases in sugarcane and effects of population substructure.

Association between markers and sugarcane diseases were investigated in a collection of 154 sugarcane clones, consisting of important ancestors or parents, and cultivars. 1,068 polymorphic AFLP and 141 SRR markers were scored across all clones. Data on the four most important diseases in the Australian sugarcane industry were obtained; these diseases being pachymetra root rot (Pachymetra chaunorhiza B.J. Croft & M.W. Dick), leaf scald (Xanthomonas albilineans Dowson), Fiji leaf gall (Fiji disease virus), and smut (Ustilago scitaminea H. & P. Sydow). By a simple regression analysis, association between markers and diseases could be readily detected. However, many of these associations were due to the effects of embedded population structure and random effects. After taking population structure into account, we found that 59% of the phenotypic variation in smut resistance ratings could be accounted for by 11 markers, 32% of variation for leaf scald and pachymetra root rot rating by 4 markers, and 26% of Fiji leaf gall by 5 markers. The results suggest that marker-trait associations can be readily detected in populations generated from modern sugarcane breeding programs. This may be due to special features of past sugarcane breeding programs leading to persistent linkage disequilibrium in modern parental populations.

Genetic Linkage↗