PubMed Health⌕ Search

Biomedical subjects

Rajeev K Varshney

Publications and source records attributed to Rajeev K Varshney.

11 recordsLinked to original sources

Pan-genome-based resequencing of 2,320 accessions reveals structural variations and accelerates breeding advances in cultivated peanut.

The cultivated peanut is a crucial global legume crop that is essential for food security and nutrition, particularly in developing regions. However, its limited genetic variation hampers breeding progress and yield improvement. Here we constructed a graph-based pan-genome for peanut, incorporating 14 genomes that represent all 6 peanut varieties. Using this pan-genome, we genotyped 2,320 accessions, covering 88.03% of ICRISAT and 59.21% of USDA core germplasm, enriching valuable resources for genomic studies and breeding. We cataloged genomic structural variations and investigated the role of homoeologous exchanges in population divergence. Through our pan-genome approach, we overcame the challenges of genotyping posed by homoeologous exchanges and identified key genes associated with flowering and dwarfism in peanut. By integrating superior haplotypes and germplasm resources guided by the pan-genome, we further developed high-yield dwarf lines. This work provides essential genomic resources to accelerate functional gene discovery and modern peanut breeding.

Journal Article↗

Functional analysis of a GWAS pleiotropic hotspot suggests an auxin biosynthesis gene (AhPDS1), regulating pod development in peanut (Arachis hypogaea L.).

Peanut productivity and quality improvement rely on understanding the genetic factors influencing pod and seed size. This study aims to identify genetic factors and regulatory mechanisms influencing pod and seed size in peanuts. Herein, a genome-wide association study (GWAS) was conducted using 390 accessions from 15 peanut growing regions to analyze pod and seed traits across multiple planting seasons. A significant phenotypic variation was observed, with broad-sense heritability ranging from 53.6 to 85.4%. Strong correlations between pod and seed traits further suggest potential for co-selection in breeding efforts. A pleiotropic hotspot on chromosome B06 was strongly associated with six pod and seed traits. A peanut pod size regulator AhPDS1 (PODSIZE-1, Ahy_B06g085516) homolog of Arabidopsis thaliana YUCCA4 (AtYUC4, AT5G11320), involved in auxin biosynthesis, was selected as a candidate regulating pod and seed size. Quantitative reverse transcriptase-polymerase chain reaction (qRT-PCR) confirmed higher AhPDS1 expression in large pod as compared with the small pod genotypes. Subcellular localization showed AhPDS1 to be predominantly cytoplasmic, and GUS reporter assays indicated widespread expression in roots, stems, leaves, flowers, and pods, suggesting a broad functional role. Further overexpression of AhPDS1 in Arabidopsis and rice enhanced pod, seed, and grain sizes via the indole-3-pyruvic acid pathway in transgene lines. These findings highlight AhPDS1 as a potential target for peanut molecular breeding, offering opportunities to enhance pod size via auxin biosynthesis and support sustainable crop improvement.

Arachis↗

A 1,000-loci transcript map of the barley genome: new anchoring points for integrative grass genomics.

An integrated barley transcript map (consensus map) comprising 1,032 expressed sequence tag (EST)-based markers (total 1,055 loci: 607 RFLP, 190 SSR, and 258 SNP), and 200 anchor markers from previously published data, has been generated by mapping in three doubled haploid (DH) populations. Between 107 and 179 EST-based markers were allocated to the seven individual barley linkage groups. The map covers 1118.3 cM with individual linkage groups ranging from 130 cM (chromosome 4H) to 199 cM (chromosome 3H), yielding an average marker interval distance of 0.9 cM. 475 EST-based markers showed a syntenic organisation to known colinear linkage groups of the rice genome, providing an extended insight into the status of barley/rice genome colinearity as well as ancient genome duplications predating the divergence of rice and barley. The presented barley transcript map is a valuable resource for targeted marker saturation and identification of candidate genes at agronomically important loci. It provides new anchor points for detailed studies in comparative grass genomics and will support future attempts towards the integration of genetic and physical mapping information.

Chromosome Mapping↗

Identification, characterization and utilization of EST-derived genic microsatellite markers for genome analyses of coffee and related species.

Genic microsatellites or EST-SSRs derived from expressed sequence tags (ESTs) are desired because these are inexpensive to develop, represent transcribed genes, and often a putative function can be assigned to them. In this study we investigated 2,553 coffee ESTs (461 from the public domain and 2,092 in-house generated ESTs) for identification and development of genic microsatellite markers. Of these, 2,458 ESTs (all >100 bp in size) were searched for SSRs using MISA--search module followed by stackPACK clustering that revealed a total of 425 microsatellites in 331 (13.5%) non-redundant ESTs/consensus sequences suggesting an approximate frequency of 1 SSR/2.16 kb of the analysed coffee transcriptome. Identified microsatellites mainly comprised of di-/tri-nucleotide repeats, of which repeat motifs AG and AAG were the most abundant. A total of 224 primer pairs could be designed from the non-redundant SSR-positive ESTs (excluding those with only mononucleotide repeats) for possible use as potential genic markers. Of this set, a total of 24 (10%) primer pairs were tested and 18 could be validated as usable markers. Sixteen of these markers revealed moderate to high polymorphism information content (PIC) across 23 genotypes of C. arabica and C. canephora, while 2 markers were found to be monomorphic. All the markers also showed robust cross-species amplifications across 14 Coffea and 4 Psilanthus species. The apparent broad cross-species/genera transferability was further confirmed by cloning and sequencing of the amplified alleles. Thus, the study provides an insight about the frequency and distribution of SSRs in coffee transcriptome, and also demonstrates the successful development of genic-SSRs. It is expected that the potential markers described here would add to the repertoire of DNA markers needed for genetic studies in cultivated coffee and also related taxa that constitute the important secondary genepool for coffee improvement.

Base Sequence↗

Recent history of artificial outcrossing facilitates whole-genome association mapping in elite inbred crop varieties.

Genomewide association studies depend on the extent of linkage disequilibrium (LD), the number and distribution of markers, and the underlying structure in populations under study. Outbreeding species generally exhibit limited LD, and consequently, a very large number of markers are required for effective whole-genome association genetic scans. In contrast, several of the world's major food crops are self-fertilizing inbreeding species with narrow genetic bases and theoretically extensive LD. Together these are predicted to result in a combination of low resolution and a high frequency of spurious associations in LD-based studies. However, inbred elite plant varieties represent a unique human-induced pseudo-outbreeding population that has been subjected to strong selection for advantageous alleles. By assaying 1,524 genomewide SNPs we demonstrate that, after accounting for population substructure, the level of LD exhibited in elite northwest European barley, a typical inbred cereal crop, can be effectively exploited to map traits by using whole-genome association scans with several hundred to thousands of biallelic SNPs.

Crosses, Genetic↗

Advances in cereal genomics and applications in crop breeding.

Recent advances in cereal genomics have made it possible to analyse the architecture of cereal genomes and their expressed components, leading to an increase in our knowledge of the genes that are linked to key agronomically important traits. These studies have used molecular genetic mapping of quantitative trait loci (QTL) of several complex traits that are important in breeding. The identification and molecular cloning of genes underlying QTLs offers the possibility to examine the naturally occurring allelic variation for respective complex traits. Novel alleles, identified by functional genomics or haplotype analysis, can enrich the genetic basis of cultivated crops to improve productivity. Advances made in cereal genomics research in recent years thus offer the opportunities to enhance the prediction of phenotypes from genotypes for cereal breeding.

Alleles↗

Laboratory Information Management Software for genotyping workflows: applications in high throughput crop genotyping.

BACKGROUND: With the advances in DNA sequencer-based technologies, it has become possible to automate several steps of the genotyping process leading to increased throughput. To efficiently handle the large amounts of genotypic data generated and help with quality control, there is a strong need for a software system that can help with the tracking of samples and capture and management of data at different steps of the process. Such systems, while serving to manage the workflow precisely, also encourage good laboratory practice by standardizing protocols, recording and annotating data from every step of the workflow. RESULTS: A laboratory information management system (LIMS) has been designed and implemented at the International Crops Research Institute for the Semi-Arid Tropics (ICRISAT) that meets the requirements of a moderately high throughput molecular genotyping facility. The application is designed as modules and is simple to learn and use. The application leads the user through each step of the process from starting an experiment to the storing of output data from the genotype detection step with auto-binning of alleles; thus ensuring that every DNA sample is handled in an identical manner and all the necessary data are captured. The application keeps track of DNA samples and generated data. Data entry into the system is through the use of forms for file uploads. The LIMS provides functions to trace back to the electrophoresis gel files or sample source for any genotypic data and for repeating experiments. The LIMS is being presently used for the capture of high throughput SSR (simple-sequence repeat) genotyping data from the legume (chickpea, groundnut and pigeonpea) and cereal (sorghum and millets) crops of importance in the semi-arid tropics. CONCLUSION: A laboratory information management system is available that has been found useful in the management of microsatellite genotype data in a moderately high throughput genotyping laboratory. The application with source code is freely available for academic users and can be downloaded from http://www.icrisat.org/gt-bt/lims/lims.asp.

Algorithms↗

Genomics-assisted breeding for crop improvement.

Genomics research is generating new tools, such as functional molecular markers and informatics, as well as new knowledge about statistics and inheritance phenomena that could increase the efficiency and precision of crop improvement. In particular, the elucidation of the fundamental mechanisms of heterosis and epigenetics, and their manipulation, has great potential. Eventually, knowledge of the relative values of alleles at all loci segregating in a population could allow the breeder to design a genotype in silico and to practice whole genome selection. High costs currently limit the implementation of genomics-assisted crop improvement, particularly for inbreeding and/or minor crops. Nevertheless, marker-assisted breeding and selection will gradually evolve into 'genomics-assisted breeding' for crop improvement.

Breeding↗

Genic microsatellite markers in plants: features and applications.

Expressed sequence tag (EST) projects have generated a vast amount of publicly available sequence data from plant species; these data can be mined for simple sequence repeats (SSRs). These SSRs are useful as molecular markers because their development is inexpensive, they represent transcribed genes and a putative function can often be deduced by a homology search. Because they are derived from transcripts, they are useful for assaying the functional diversity in natural populations or germplasm collections. These markers are valuable because of their higher level of transferability to related species, and they can often be used as anchor markers for comparative mapping and evolutionary studies. They have been developed and mapped in several crop species and could prove useful for marker-assisted selection, especially when the markers reside in the genes responsible for a phenotypic trait. Applications and potential uses of EST-SSRs in plant genetics and breeding are discussed.

Chromosome Mapping↗

Large-scale analysis of the barley transcriptome based on expressed sequence tags.

To provide resources for barley genomics, 110,981 expressed sequence tags (ESTs) were generated from 22 cDNA libraries representing tissues at various developmental stages. This EST collection corresponds to approximately one-third of the 380,000 publicly available barley ESTs. Clustering and assembly resulted in 14,151 tentative consensi (TCs) and 11 073 singletons, altogether representing 25 224 putatively unique sequences. Of these, 17.5% showed no significant similarity to other barley ESTs present in dbEST. More than 41% of all barley genes are supposed to belong to multigene families and approximately 4% of the barley genes undergo alternative splicing. Based on the functional annotation of the set of unique sequences, the functional category 'Energy' was further analysed to reveal tissue- and stage-specific differences in gene expression. Hierarchical clustering of 362 differentially expressed TCs resulted in the identification of seven major clusters. The clusters reflect biochemical pathways predominantly activated in specific tissues and at various developmental stages. During seed germination glycolysis could be identified as the most predominant biochemical pathway. Germination-specific glycolysis is characterized by the coordinated expression of phosphoenolpyruvate carboxylase and phosphoenolpyruvate carboxykinase, whose antagonistic actions possibly regulate the flux of amino acids into protein biosynthesis and gluconeogenesis respectively. The expression of defence-related and antioxidant genes during germination might be controlled by the ethylene-signalling pathway as concluded from the coordinated expression of those genes and the transcription factors (TF) EIN3 and EREBPG. Moreover, because of their predominant expression in germinating seeds, TF of the AP2 and MYB type are presumably major regulators of germination.

Expressed Sequence Tags↗

In silico analysis on frequency and distribution of microsatellites in ESTs of some cereal species.

During the last decade microsatellites or SSRs (simple sequence repeats) have been proven to be the markers of choice in plant genetics research and for breeding purposes because of their hypervariability and ease of detection. However, development of these markers is expensive, labour intensive and time consuming, in particular, if they are being developed from genomic libraries. In the context of large-scale sequencing and genomics programmes in various cereal species at different laboratories, a large set of expressed sequence tags (ESTs) is being generated, which can be used to search for microsatellites. Keeping in view the importance of such type of SSRs, available ESTs of some cereal species like barley, maize, oats, rice, rye and wheat were investigated for a study of abundance, frequency and distribution of various types of microsatellites. SSRs were present in about 7% to 10% of the total ESTs in the investigated cereal genomes. On the basis of surveying EST sequences amounting to 75.2 Mb in barley, 54.7 Mb in maize, 43.9 Mb in rice, 3.7 Mb in rye, 41.6 Mb in sorghum and 37.5 Mb in wheat, the frequency of SSRs was 1/7.5 kb in barley, 1/7.5 kb in maize, 1/6.2 kb in wheat, 1/5.5 kb in rye and sorghum and 1/3.9 kb in rice. The overall average SSR frequency for these species is 1/6.0 kb. Trimeric repeats are the most abundant (54% to 78%) class of microsatellites followed by dimeric repeats (17% to 40%). Among the trimeric repeats the motifs CCG are the most common in all the cases ranging from 32% in wheat to 49% in sorghum. When all these SSRs were analysed for assessing their potential to develop new markers, unique primer pairs could be designed for 30% to 70% of the total non-redundant microsatellites which are up to 3% of total ESTs in the studied species.

Crops, Agricultural↗