Article
Beyond Accuracy: An Efficiency- and Safety-Aware Framework for Evaluating Clinical AI with Large Language Models
2025-10-19
Abstract excerpt
<h4>Background</h4> Large language models (LLMs) demonstrate strong performance on medical reasoning tasks, but current evaluation approaches focus primarily on accuracy, neglecting the efficiency–safety trade-offs critical for real-world clinical utility. <h4>Methods</h4> We developed and validated the Clinical Value Density (CVD) framework, a novel metric quantifying clinical utility per unit of cognitive reso...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 152072be-9cdb-50fc-a6cf-26b34cf0e9c9
- DOI
- 10.1101/2025.10.14.25338039
