BAC-pool sequencing and analysis of large segments of A12 and D12 homoeologous chromosomes in upland cotton.
View/ Open
Date
2013-10-08Author
Buyyarapu, R
Kantety, RV
Yu, JZ
Xu, Z
Kohel, RJ
Percy, RG
Macmil, S
Wiley, GB
Roe, Bruce A.
Sharma, GC
Metadata
Show full item recordAbstract
Although new and emerging next-generation sequencing (NGS) technologies have reduced sequencing costs significantly, much work remains to implement them for de novo sequencing of complex and highly repetitive genomes such as the tetraploid genome of Upland cotton (Gossypium hirsutum L.). Herein we report the results from implementing a novel, hybrid Sanger/454-based BAC-pool sequencing strategy using minimum tiling path (MTP) BACs from Ctg-3301 and Ctg-465, two large genomic segments in A12 and D12 homoeologous chromosomes (Ctg). To enable generation of longer contig sequences in assembly, we implemented a hybrid assembly method to process ~35x data from 454 technology and 2.8-3x data from Sanger method. Hybrid assemblies offered higher sequence coverage and better sequence assemblies. Homology studies revealed the presence of retrotransposon regions like Copia and Gypsy elements in these contigs and also helped in identifying new genomic SSRs. Unigenes were anchored to the sequences in Ctg-3301 and Ctg-465 to support the physical map. Gene density, gene structure and protein sequence information derived from protein prediction programs were used to obtain the functional annotation of these genes. Comparative analysis of both contigs with Arabidopsis genome exhibited synteny and microcollinearity with a conserved gene order in both genomes. This study provides insight about use of MTP-based BAC-pool sequencing approach for sequencing complex polyploid genomes with limited constraints in generating better sequence assemblies to build reference scaffold sequences. Combining the utilities of MTP-based BAC-pool sequencing with current longer and short read NGS technologies in multiplexed format would provide a new direction to cost-effectively and precisely sequence complex plant genomes.
Collections
Related items
Showing items related by title, author, creator and subject.
-
Novel genomes and genome constitutions identified by GISH and 5S rDNA and knotted1 genomic sequences in the genus Setaria
Zhao, Meicheng; Zhi, Hui; Doust, Andrew N.; Li, Wei; Wang, Yongfang; Li, Haiquan; Jia, Guanqing; Wang, Yongqiang; Zhang, Ning; Diao, Xianmin (BioMed Central, 2013-04-11)Background: The Setaria genus is increasingly of interest to researchers, as its two species, S. viridis and S. italica, are being developed as models for understanding C4 photosynthesis and plant functional genomics. The ... -
Building a model: Developing genomic resources for common milkweed (Asclepias syriaca) with low coverage genome sequencing
Straub, Shannon C. K.; Fishbein, Mark; Livshultz, Tatyana; Foster, Zachary; Parks, Matthew; Weitemier, Kevin; Cronn, Richard C.; Liston, Aaron (BioMed Central, 2011-05-04)Background: Milkweeds (Asclepias L.) have been extensively investigated in diverse areas of evolutionary biology and ecology; however, there are few genetic resources available to facilitate and compliment these studies. ... -
Comparison of multiple antibiotic resistant Staphylococcus aureus genomes and the genome structure of Elizabethkingia meningoseptica
Matyi, Stephanie Ann (2014-05)i. The aim of this study was to determine if methicillin-resistant Staphylococcus aureus (MRSA) strains could be identified in the milk of dairy cattle in a Paso Del Norte region dairy. Using physiological and PCR-based ...