Back to search

Article

Adding layers of information to scRNA-seq data using pre-trained language models

2025-08-27

Abstract excerpt

Pre-trained language models promise to enrich analyses of single-cell data with additional layers of information leveraging large text corpora. Yet, it is still unclear how to achieve optimal alignment with the primary quantitative single-cell data. To address this, we construct text-based training datasets from both scRNA-seq data and biomedical literature targeted to the experimental setting at hand. We then joi...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
375fc8a6-5b84-51c6-be3f-3ff6c554062b
DOI
10.1101/2025.08.23.671699
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Adding layers of information to scRNA-seq data using pre-trained language modelsDOI 10.1101/2025.08.23.671699
Select a neighboring publication to make it the new centre.