Article
Domain-specific embeddings uncover latent genetics knowledge
2025-03-19
Abstract excerpt
The inundating rate of scientific publishing means every researcher will miss new discoveries from overwhelming saturation. To address this limitation, we employ natural language processing to overcome human limitations in reading, curation, and knowledge synthesis, with domain-specific applications to genetics and genomics. We construct a corpus of 3.5 million normalized genetics and genomics abstracts and implem...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- a5f8c4ca-5789-5bdd-9022-78cd27625703
- DOI
- 10.1101/2025.03.17.643817
