PubMed Health⌕ Search

Biomedical subjects

Xiaokang Pan

Publications and source records attributed to Xiaokang Pan.

5 recordsLinked to original sources

SynBrowse: a synteny browser for comparative sequence analysis.

MOTIVATION: The recent efforts of various sequence projects to sequence deeply into various phylogenies provide great resources for comparative sequence analysis. A generic and portable tool is essential for scientists to visualize and analyze sequence comparisons. RESULTS: We have developed SynBrowse, a synteny browser for visualizing and analyzing genome alignments both within and between species. It is intended to help scientists study macrosynteny, microsynteny and homologous genes between sequences. It can also aid with the identification of uncharacterized genes, putative regulatory elements and novel structural features of a species. SynBrowse is a GBrowse (the Generic Genome Browser) family software tool that runs on top of the open source BioPerl modules. It consists of two components: a web-based front end and a set of relational database back ends. Each database stores pre-computed alignments from a focus sequence to reference sequences in addition to the genome annotations of the focus sequence. The user interface lets end users select a key comparative alignment type and search for syntenic blocks between two sequences and zoom in to view the relationships among the corresponding genome annotations in detail. SynBrowse is portable with simple installation, flexible configuration, convenient data input and easy integration with other components of a model organism system. AVAILABILITY: The software is available at http://www.gmod.org CONTACT: vbrendel@iastate.edu

Algorithms↗

Site preferences of insertional mutagenesis agents in Arabidopsis.

We have performed a comparative analysis of the insertion sites of engineered Arabidopsis (Arabidopsis thaliana) insertional mutagenesis vectors that are based on the maize (Zea mays) transposable elements and Agrobacterium T-DNA. The transposon-based agents show marked preference for high GC content, whereas the T-DNA-based agents show preference for low GC content regions. The transposon-based agents show a bias toward insertions near the translation start codons of genes, while the T-DNAs show a predilection for the putative transcriptional regulatory regions of genes. The transposon-based agents also have higher insertion site densities in exons than do the T-DNA insertions. These observations show that the transposon-based and T-DNA-based mutagenesis techniques could complement one another well, and neither alone is sufficient to achieve the goal of saturation mutagenesis in Arabidopsis. These results also suggest that transposon-based mutagenesis techniques may prove the most effective for obtaining gene disruptions and for generating gene traps, while T-DNA-based agents may be more effective for activation tagging and enhancer trapping. From the patterns of insertion site distributions, we have identified a set of nucleotide sequence motifs that are overrepresented at the transposon insertion sites. These motifs may play a role in the transposon insertion site preferences. These results could help biologists to study the mechanisms of insertions of the insertional mutagenesis agents and to design better strategies for genome-wide insertional mutagenesis.

Arabidopsis↗

Maize-targeted mutagenesis: A knockout resource for maize.

We describe an efficient system for site-selected transposon mutagenesis in maize. A total of 43,776 F1 plants were generated by using Robertson's Mutator (Mu) pollen parents and self-pollinated to establish a library of transposon-mutagenized seed. The frequency of new seed mutants was between 10-4 and 10-5 per F1 plant. As a service to the maize community, maize-targeted mutagenesis selects insertions in genes of interest from this library by using the PCR. Pedigree, knockout, sequence, phenotype, and other information is stored in a powerful interactive database (maize-targeted mutagenesis database) that enables analysis of the entire population and the handling of knockout requests. By inhibiting Mu activity in most F1 plants, we sought to reduce somatic insertions that may cause false positives selected from pooled tissue. By monitoring the remaining Mu activity in the F2, however, we demonstrate that seed phenotypes depend on it, and false positives occur in lines that appear to lack it. We conclude that more than half of all mutations arising in this population are suppressed on losing Mu activity. These results have implications for epigenetic models of inbreeding and for functional genomics.

Base Sequence↗

ATIDB: Arabidopsis thaliana insertion database.

Insertional mutagenesis techniques, including transposon- and T-DNA-mediated mutagenesis, are key resources for systematic identification of gene function in the model plant species Arabidopsis thaliana. We have developed a database (http://atidb.cshl.org/) for archiving, searching and analyzing insertional mutagenesis lines. Flanking sequences from approximately 10 500 insertion lines (including transposon and T-DNA insertions) from several tagging programs in Arabidopsis were mapped to the genome sequence through our annotation system before being entered into the database. The database front end provides World Wide Web searching and analyzing interfaces for genome researchers and other biologists. Users can search the database to identify insertions in a particular gene or perform genome-wide analysis to study the distribution and preference of insertions. Tools integrated with the database include a graphical genome browser, a protein search function, a graphical representation of the insertion distribution and a Blast search function. The database is based on open source components and is available under an open source license.

Amino Acid Sequence↗

Gramene: a resource for comparative grass genomics.

Gramene (http://www.gramene.org) is a comparative genome mapping database for grasses and a community resource for rice. Rice, in addition to being an economically important crop, is also a model monocot for understanding other agronomically important grass genomes. Gramene replaces the existing AceDB database 'RiceGenes' with a relational database based on Oracle. Gramene provides curated and integrative information about maps, sequence, genes, genetic markers, mutants, QTLs, controlled vocabularies and publications. Its aims are to use the rice genetic, physical and sequence maps as fundamental organizing units, to provide a common denominator for moving from one crop grass to another and is to serve as a portal for interconnecting with other web-based crop grass resources. This paper describes the initial steps we have taken towards realizing these goals.

Chromosome Mapping↗