PubMed Health⌕ Search

PubMed · 16931491

A genome-wide study of preferential amplification/hybridization in microarray-based pooled DNA experiments.

Abstract

Microarray-based pooled DNA methods overcome the cost bottleneck of simultaneously genotyping more than 100 000 markers for numerous study individuals. The success of such methods relies on the proper adjustment of preferential amplification/hybridization to ensure accurate and reliable allele frequency estimation. We performed a hybridization-based genome-wide single nucleotide polymorphisms (SNPs) genotyping analysis to dissect preferential amplification/hybridization. The majority of SNPs had less than 2-fold signal amplification or suppression, and the lognormal distributions adequately modeled preferential amplification/hybridization across the human genome. Comparative analyses suggested that the distributions of preferential amplification/hybridization differed among genotypes and the GC content. Patterns among different ethnic populations were similar; nevertheless, there were striking differences for a small proportion of SNPs, and a slight ethnic heterogeneity was observed. To fulfill appropriate and gratuitous adjustments, databases of preferential amplification/hybridization for African Americans, Caucasians and Asians were constructed based on the Affymetrix GeneChip Human Mapping 100 K Set. The robustness of allele frequency estimation using this database was validated by a pooled DNA experiment. This study provides a genome-wide investigation of preferential amplification/hybridization and suggests guidance for the reliable use of the database. Our results constitute an objective foundation for theoretical development of preferential amplification/hybridization and provide important information for future pooled DNA analyses.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

H-C Yang, Y-J Liang, M-C Huang, L-H Li, C-H Lin, J-Y Wu, Y-T Chen, C S J Fann. 2006-08-23. A genome-wide study of preferential amplification/hybridization in microarray-based pooled DNA experiments.. https://doi.org/10.1093/nar%2Fgkl446

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

The INSDC specifications-foundations for a FAIR and global INSDC.

Members of the International Nucleotide Sequence Database Collaboration (INSDC; https://www.insdc.org/) collect, exchange, and preserve comprehensive open nucleotide sequence information and provide tools for its access. The INSDC has stated its commitment to welcoming new members into the collaboration to be more representative of the global community of data and users. To reach this goal, a comprehensive definition of the INSDC data model and minimum requirements for data acceptance have been established. Here we describe the processes used to arrive upon these INSDC Specifications and lay out strategies for their continued upkeep to remain current and relevant. Database URL:  https://www.insdc.org/.

Databases, Nucleic Acid↗

Pyrosequencing genotype storage techniques.

Data storage and data coordination are important aspects of project design and execution. Pyrosequencing technology allows thousands of data-points to be collected per day. Consequently, a consistent and reliable method of data input and storage is vital. This chapter discusses the strengths and weaknesses of data storage systems.

Databases, Nucleic Acid↗

CSRDB: a small RNA integrated database and browser resource for cereals.

Plant small RNAs (smRNAs), which include microRNAs (miRNAs), short interfering RNAs (siRNAs) and trans-acting siRNAs (ta-siRNAs), are emerging as significant components of epigenetic processes and of gene networks involved in development and in homeostasis. Here we present a bioinformatics resource for cereal crops, the Cereal Small RNA Database (CSRDB), consisting of large-scale datasets of maize and rice smRNA sequences generated by high-throughput pyrosequencing. The smRNA sequences have been mapped to the rice genome and to the available maize genome sequence and these results are presented in two genome browser datasets using the Generic Genome Browser. Potential RNA targets for the smRNAs have been predicted and access to the resulting smRNA/RNA target pair dataset has been made available through a MySQL based relational database. Various ways to access the data are provided including links from the genome browser to the target database. Data linking and integration are the main focus for this interface, and internal as well as external links are present. The resource is available at http://sundarlab.ucdavis.edu/smrnas/ and will be updated as more sequences become available.

Databases, Nucleic Acid↗