Article
Using reference-free compressed data structures to analyse sequencing reads from thousands of human genomes
2016-06-22
Abstract excerpt
We are rapidly approaching the point where we have sequenced millions of human genomes. There is a pressing need for new data structures to store raw sequencing data and efficient algorithms for population scale analysis. Current reference based data formats do not fully exploit the redundancy in population sequencing nor take advantage of shared genetic variation. In recent years, the Burrows-Wheeler transform (B...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 3c5d6393-0365-5ee5-8275-1d9d3ed8fe6b
- DOI
- 10.1101/060186
