Article
Confidence Is Not Probability: Verification-Based Confidence Prediction for LLM-Generated Clinical Summaries
2026-07-28
Abstract excerpt
<h4>Objective: </h4> Token-level confidence estimates from large language models (LLMs) do not reflect whether a generated clinical summary is factually correct. We investigate whether verification-based confidence prediction improves upon token-level estimates and whether LLM-as-judge labels can be assumed reliable for clinical summarization evaluation. Methods. Using MIMIC-IV, we construct 500 clinical summariza...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- b8cf0f75-79f2-5158-809e-32a9ff9901c6
- DOI
- 10.20944/preprints202607.1999.v1
