Article
A bioinformatics approach for determining sample identity from different lanes of high-throughput sequencing data.
PloS one - 1 Jan 2011
Goldfeder Rachel L, Parker Stephen C J, Ajay Subramanian S, Ozel Abaan Hatice, Margulies Elliott H
Abstract excerpt
The ability to generate whole genome data is rapidly becoming commoditized. For example, a mammalian sized genome (∼3Gb) can now be sequenced using approximately ten lanes on an Illumina HiSeq 2000. Since lanes from different runs are often combined, verifying that each lane in a genome's build is from the same sample is an important quality control. We sought to address this issue in a post hoc bioinformatic...
Read the complete abstract on PubMedTopics
Share this publication in a Topic to start or enrich a Post.
