Back to search

Article

Embarrassingly_FASTA: Enabling Recomputable, Population-Scale Pangenomics by Reducing Commercial Genome Processing Costs from $100 to less than $1

2026-02-04

Abstract excerpt

Computational preprocessing has become the dominant bottleneck in genomics, frequently exceeding sequencing costs and constraining population-scale analysis, even as large repositories grow from tens of petabytes toward exabyte-scale storage to support World Genome Models. Legacy CPU-based workflows require many hours to days per 30× human genome, driving many repositories to distribute aligned or derived intermed...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
dc350798-0ef5-5cdf-837a-c30aa9eb226a
DOI
10.64898/2026.02.02.703356
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Embarrassingly_FASTA: Enabling Recomputable, Population-Scale Pangenomics by Reducing Commercial Genome Processing Costs from $100 to less than $1DOI 10.64898/2026.02.02.703356
Select a neighboring publication to make it the new centre.