Whole-genome sequencing and variant discovery in C. elegans |
| |
Authors: | Hillier LaDeana W Marth Gabor T Quinlan Aaron R Dooling David Fewell Ginger Barnett Derek Fox Paul Glasscock Jarret I Hickenbotham Matthew Huang Weichun Magrini Vincent J Richt Ryan J Sander Sacha N Stewart Donald A Stromberg Michael Tsung Eric F Wylie Todd Schedl Tim Wilson Richard K Mardis Elaine R |
| |
Affiliation: | Washington University School of Medicine, Department of Genetics and Genome Sequencing Center, 4444 Forest Park Blvd., St. Louis, Missouri 63108, USA. |
| |
Abstract: | Massively parallel sequencing instruments enable rapid and inexpensive DNA sequence data production. Because these instruments are new, their data require characterization with respect to accuracy and utility. To address this, we sequenced a Caernohabditis elegans N2 Bristol strain isolate using the Solexa Sequence Analyzer, and compared the reads to the reference genome to characterize the data and to evaluate coverage and representation. Massively parallel sequencing facilitates strain-to-reference comparison for genome-wide sequence variant discovery. Owing to the short-read-length sequences produced, we developed a revised approach to determine the regions of the genome to which short reads could be uniquely mapped. We then aligned Solexa reads from C. elegans strain CB4858 to the reference, and screened for single-nucleotide polymorphisms (SNPs) and small indels. This study demonstrates the utility of massively parallel short read sequencing for whole genome resequencing and for accurate discovery of genome-wide polymorphisms. |
| |
Keywords: | |
本文献已被 PubMed 等数据库收录! |
|