Back to search

Article

Using reference-free compressed data structures to analyse sequencing reads from thousands of human genomes

2016-06-22

Abstract excerpt

We are rapidly approaching the point where we have sequenced millions of human genomes. There is a pressing need for new data structures to store raw sequencing data and efficient algorithms for population scale analysis. Current reference based data formats do not fully exploit the redundancy in population sequencing nor take advantage of shared genetic variation. In recent years, the Burrows-Wheeler transform (B...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
3c5d6393-0365-5ee5-8275-1d9d3ed8fe6b
DOI
10.1101/060186
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Using reference-free compressed data structures to analyse sequencing reads from thousands of human genomesDOI 10.1101/060186
Select a neighboring publication to make it the new centre.