PubMed HealthSearch

PubMed · 8787973

Data simulation for GAW9 problems 1 and 2.

Abstract

Herein we describe the methods utilized to simulate the genetic marker data for GAW9 Problems 1 and 2, as well as the pedigree and phenotype data for GAW9 Problem 1.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

M C Speer, J D Terwilliger, J Ott. 1995. Data simulation for GAW9 problems 1 and 2.. https://doi.org/10.1002/gepi.1370120606

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Generating correlated data for omics simulation.

Simulation of realistic omics data is a key input for benchmarking studies that help users obtain optimal computational pipelines. Omics data involves large numbers of measured features on each sample and these measures are generally correlated with each other. However, simulation too often ignores these correlations, perhaps due to computational and statistical hurdles of doing so. To alleviate this, we describe three approaches for generating omics-scale data with correlated measures which mimic real datasets. These approaches are all based on a Gaussian copula approach with a covariance matrix that decomposes into a diagonal part and a low-rank part. This decomposition allows for extremely efficient simulation, overcoming a hurdle for adoption of past methods. We use these approaches to demonstrate the importance of including correlation in two benchmarking applications. First, we show that variance of results from the popular DESeq2 method increases when dependence is included. Second, we demonstrate that CYCLOPS, a method for inferring circadian time of collection from transcriptomics, improves in performance when given gene-gene dependencies in some circumstances. We provide an R package, dependentsimr, that has efficient implementations of these methods and can generate dependent data with arbitrary marginal distributions, including discrete (binary, ordered categorical, Poisson, negative binomial), continuous (normal), or with an empirical distribution.

Computer Simulation

Addressing current challenges in cancer immunotherapy with mathematical and computational modelling.

The goal of cancer immunotherapy is to boost a patient's immune response to a tumour. Yet, the design of an effective immunotherapy is complicated by various factors, including a potentially immunosuppressive tumour microenvironment, immune-modulating effects of conventional treatments and therapy-related toxicities. These complexities can be incorporated into mathematical and computational models of cancer immunotherapy that can then be used to aid in rational therapy design. In this review, we survey modelling approaches under the umbrella of the major challenges facing immunotherapy development, which encompass tumour classification, optimal treatment scheduling and combination therapy design. Although overlapping, each challenge has presented unique opportunities for modellers to make contributions using analytical and numerical analysis of model outcomes, as well as optimization algorithms. We discuss several examples of models that have grown in complexity as more biological information has become available, showcasing how model development is a dynamic process interlinked with the rapid advances in tumour-immune biology. We conclude the review with recommendations for modellers both with respect to methodology and biological direction that might help keep modellers at the forefront of cancer immunotherapy development.

Computer Simulation

Modeling high-resolution hydration patterns in correlation with DNA sequence and conformation.

Hydration around the DNA fragment d(C5T5).(A5G5) is presented from two molecular dynamics simulations of 10 and 12 ns total simulation time. The DNA has been simulated as a flexible molecule with both the CHARMM and AMBER force fields in explicit solvent including counterions and 0.8 M additional NaCl salt. From the previous analysis of the DNA structure B-DNA conformations were found with the AMBER force-field and A-DNA conformations with CHARMM parameters. High-resolution hydration patterns are compared between the two conformations and between C.G and T.A base-pairs from the homopolymeric parts of the simulated sequence. Crystallographic results from a statistical analysis of hydration sites around DNA crystal structures compare very well with the simulation results. Differences between the crystal sites and our data are explained by variations in conformation, sequence, and limitations in the resolution of water sites by crystal diffraction. Hydration layers are defined from radial distribution functions and compared with experimental results. Excellent agreement is found when the measured experimental quantities are compared with the equivalent distribution of water molecules in the first hydration shell. The number of water molecules bound to DNA was found smaller around T.A base-pairs and around A-DNA as compared to B-DNA. This is partially offset by a larger number of water molecules in hydrophobic contact with DNA around T.A base-pairs and around A-DNA. The numbers of water molecules in minor and major grooves have been correlated with helical roll, twist, and inclination angles. The data more fully explain the observed B-->A transition at low humidity.

Computer Simulation