Article
A Clinical Speech Corpus with Temporally Aligned Sensitive Health Information
2026-04-03
Abstract excerpt
<h4>ABSTRACT</h4> <h4>Objectives</h4> Due to privacy constraints, the sensitivity of medical speech, and the complexity of speech-level annotation, publicly available datasets for clinical speech de-identification remain scarce. To address this gap, we constructed the SREDH-AI Cup Sensitive Health Information (SHI) speech corpus, a time-aligned clinical speech dataset annotated for 26 SHI categories. <h4>Methods...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- cb20db9c-e818-50af-93b2-5f703bc359c4
- DOI
- 10.64898/2026.03.31.26349906
