Back to search

Article

TabSyM: A Generative Pipeline for Small Multi-Cohort Omics Tabular Data

2025-07-18

Abstract excerpt

Machine learning applications in biomedicine such as omics data analysis are frequently hindered by datasets that are small, high-dimensional, and affected by batch effects across different patient cohorts. To address these challenges, we introduce TabSyM, a modular generative pipeline that synthesizes high-quality, task-relevant data to improve predictive modeling. TabSyM integrates three key stages: it extends a...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
17610d31-f181-58e7-a4ff-9d5dac9cab75
DOI
10.1101/2025.07.14.664738
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
TabSyM: A Generative Pipeline for Small Multi-Cohort Omics Tabular DataDOI 10.1101/2025.07.14.664738
Select a neighboring publication to make it the new centre.