PubMed Health⌕ Search

Biomedical subjects

Najma Khalid

Publications and source records attributed to Najma Khalid.

3 recordsLinked to original sources

Single-stranded linear amplification protocol results in reproducible and reliable microarray data from nanogram amounts of starting RNA.

The range of scientific questions utilizing DNA microarray techniques is limited by the fact that these methods require 5-40 microg of high-quality total RNA. Thus, methods that reliably amplify the starting RNA amount could expand the applicability of DNA microarray technology. We developed a single-stranded linear amplification protocol (SLAP) that combines the reproducibility of in vitro transcription and the amplification robustness of polymerase chain reactions. We compared SLAP to the NIH-IVT amplification protocol. SLAP displayed excellent conservation of the 5'/3' signal and demonstrated the most robust amplification, producing the recommended amounts of biotin-labeled RNA with as little as 0.002 microg of starting RNA. Both SLAP and NIH-IVT methods demonstrated good reproducibility, but SLAP maintained the highest level of reliability with RNA starting amounts of <0.05 microg. These results suggest that SLAP is an excellent alternative to IVT-based amplification protocols when RNA is limited by small sample size.

Biotin↗

A method for the assessment of disease associations with single-nucleotide polymorphism haplotypes and environmental variables in case-control studies.

The rough draft of the human genome map has been used to identify most of the functional genes in the human genome, as well as to identify nucleotide variations, known as "single-nucleotide polymorphisms" (SNPs), in these genes. By use of advanced biotechnologies, researchers are beginning to genotype thousands of SNPs from biological samples. Among the many possible applications, one of them is the study of SNP associations with complex human diseases, such as cancers or coronary heart diseases, by using a case-control study design. Through the gathering of environmental risk factors and other lifestyle factors, such a study can be effectively used to investigate interactions between genes and environmental factors in their associations with disease phenotype. Earlier, we developed a method to statistically construct individuals' haplotypes and to estimate the distribution of haplotypes of multiple SNPs in a defined population, by use of estimating-equation techniques. Extending this idea, we describe here an analytic method for assessing the association between the constructed haplotypes along with environmental factors and the disease phenotype. This method is also robust to the model assumptions and is scalable to a large number of SNPs. Asymptotic properties of estimations in the method are proved theoretically and are tested for finite sample sizes by use of simulations. To demonstrate the use of the method, we applied it to assess the possible association between apolipoprotein CIII (six coding SNPs) and restenosis by using a case-control data set. Our analysis revealed two haplotypes that may reduce the risk of restenosis.

Apolipoprotein C-III↗

Estimating haplotype frequencies and standard errors for multiple single nucleotide polymorphisms.

Estimating haplotype frequencies becomes increasingly important in the mapping of complex disease genes, as millions of single nucleotide polymorphisms (SNPs) are being identified and genotyped. When genotypes at multiple SNP loci are gathered from unrelated individuals, haplotype frequencies can be accurately estimated using expectation-maximization (EM) algorithms (Excoffier and Slatkin, 1995; Hawley and Kidd, 1995; Long et al., 1995), with standard errors estimated using bootstraps. However, because the number of possible haplotypes increases exponentially with the number of SNPs, handling data with a large number of SNPs poses a computational challenge for the EM methods and for other haplotype inference methods. To solve this problem, Niu and colleagues, in their Bayesian haplotype inference paper (Niu et al., 2002), introduced a computational algorithm called progressive ligation (PL). But their Bayesian method has a limitation on the number of subjects (no more than 100 subjects in the current implementation of the method). In this paper, we propose a new method in which we use the same likelihood formulation as in Excoffier and Slatkin's EM algorithm and apply the estimating equation idea and the PL computational algorithm with some modifications. Our proposed method can handle data sets with large number of SNPs as well as large numbers of subjects. Simultaneously, our method estimates standard errors efficiently, using the sandwich-estimate from the estimating equation, rather than the bootstrap method. Additionally, our method admits missing data and produces valid estimates of parameters and their standard errors under the assumption that the missing genotypes are missing at random in the sense defined by Rubin (1976).

Algorithms↗