PubMed Health⌕ Search

PubMed · 17033957

Generalized genomic distance-based regression methodology for multilocus association analysis.

Abstract

Large-scale, multilocus genetic association studies require powerful and appropriate statistical-analysis tools that are designed to relate genotype and haplotype information to phenotypes of interest. Many analysis approaches consider relating allelic, haplotypic, or genotypic information to a trait through use of extensions of traditional analysis techniques, such as contingency-table analysis, regression methods, and analysis-of-variance techniques. In this work, we consider a complementary approach that involves the characterization and measurement of the similarity and dissimilarity of the allelic composition of a set of individuals' diploid genomes at multiple loci in the regions of interest. We describe a regression method that can be used to relate variation in the measure of genomic dissimilarity (or "distance") among a set of individuals to variation in their trait values. Weighting factors associated with functional or evolutionary conservation information of the loci can be used in the assessment of similarity. The proposed method is very flexible and is easily extended to complex multilocus-analysis settings involving covariates. In addition, the proposed method actually encompasses both single-locus and haplotype-phylogeny analysis methods, which are two of the most widely used approaches in genetic association analysis. We showcase the method with data described in the literature. Ultimately, our method is appropriate for high-dimensional genomic data and anticipates an era when cost-effective exhaustive DNA sequence data can be obtained for a large number of individuals, over and above genotype information focused on a few well-chosen loci.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Jennifer Wessel, Nicholas J Schork. 2006-09-21. Generalized genomic distance-based regression methodology for multilocus association analysis.. https://doi.org/10.1086/508346

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Genome sequence data of the chitinase-producing bacterium Paenibacillus mucilaginosus YWY-5.1.

Paenibacillus mucilaginosus is a beneficial bacterium widely applied as a biofertilizer in agriculture. To date, genomic information on this species remains limited; however, no genome assemblies from Vietnam have been reported. This work presented the draft genome of P. mucilaginosus YWY-5.1, a promising strain with strong chitin-degrading capability and agricultural potential, isolated from Yok Don National Park, Vietnam, using Illumina technology. Results showed that the assembled genome comprised 48 contigs with 4,076,146 bp and 73.8% GC-content. Genome annotation identified 3,611 protein-coding genes, 2 rRNA genes, and 53 tRNA genes. A total of 150 carbohydrate-active enzyme-related genes were predicted from the genome; among them, seven putative chitinolytic genes were identified, including 4 genes related to family 18 chitinase, 2 genes to family 20 β-N-acetylglucosaminidase, and one gene to auxiliary activity family 10. In addition, at least 32 genes related to plant growth-promoting functions were identified, including those associated with indole-3-acetic acid production, phosphate and potassium solubilization, siderophore biosynthesis, iron uptake, ACC metabolism, and nitrate transport and reduction. Furthermore, genome mining identified 4 biosynthetic gene clusters probably involved in secondary metabolite production, of which 3 displayed no similarity to previously reported clusters, indicating potential for novel bioactive compounds. These genomic data improved our understanding of the biodegradation capacity and agricultural potential of P. mucilaginosus YWY-5.1 isolated from Vietnam, and provided a valuable genomic resource for future functional and biotechnological investigations toward crop production and related fields.

Chitinases↗

Costs and benefits of processivity in enzymatic degradation of recalcitrant polysaccharides.

Many enzymes that hydrolyze insoluble crystalline polysaccharides such as cellulose and chitin guide detached single-polymer chains through long and deep active-site clefts, leading to processive (stepwise) degradation of the polysaccharide. We have studied the links between enzyme efficiency and processivity by analyzing the effects of mutating aromatic residues in the substrate-binding groove of a processive chitobiohydrolase, chitinase B from Serratia marcescens. Mutation of two tryptophan residues (Trp-97 and Trp-220) close to the catalytic center (subsites +1 and +2) led to reduced processivity and a reduced ability to degrade crystalline chitin, suggesting that these two properties are linked. Most remarkably, the loss of processivity in the W97A mutant was accompanied by a 29-fold increase in the degradation rate for single-polymer chains as present in the soluble chitin-derivative chitosan. The properties of the W220A mutant showed a similar trend, although mutational effects were less dramatic. Processivity is thought to contribute to the degradation of crystalline polysaccharides because detached single-polymer chains are kept from reassociating with the solid material. The present results show that this processivity comes at a large cost in terms of enzyme speed. Thus, in some cases, it might be better to focus strategies for enzymatic depolymerization of polysaccharide biomass on improving substrate accessibility for nonprocessive enzymes rather than on improving the properties of processive enzymes.

Chitinases↗

Human CHIT1 gene distribution: new data from Mediterranean and European populations.

A 24 bp duplication in the CHIT1 gene (H allele) is associated with a deficiency in the activity of chitotriosidase, an enzyme with the capability to hydrolyse chitin. A recent study in European and two sub-Saharan populations suggested a relationship between the presence of the mutation, improved environmental conditions, and the disappearance of parasitic diseases, including Plasmodium falciparum malaria. This result was not supported by the high frequency of the 24 bp duplication in a sample from Taiwan, an area with high malaria endemicity until 40 years ago. In this study, we analysed the frequency variability of the H allele in Mediterranean populations and its internal variability in Sardinia (Italy) with respect to malaria, which had been endemic on the island until its eradication during 1946-1950. The pattern of H frequency distributions is not consistent with the hypothesis of selective pressures acting on CHIT1 gene. The Moran's index coefficient and correlogram seem to indicate, indeed, that allele distribution was determined by random factors. The pattern of frequency distribution suggests a possible Asiatic origin of the H allele, but it could be possible also that the mutant allele had diffused out of Africa, and was subsequently lost from African populations.

Chitinases↗