Back to search

Article

Beyond Accuracy: An Efficiency- and Safety-Aware Framework for Evaluating Clinical AI with Large Language Models

2025-10-19

Abstract excerpt

<h4>Background</h4> Large language models (LLMs) demonstrate strong performance on medical reasoning tasks, but current evaluation approaches focus primarily on accuracy, neglecting the efficiency–safety trade-offs critical for real-world clinical utility. <h4>Methods</h4> We developed and validated the Clinical Value Density (CVD) framework, a novel metric quantifying clinical utility per unit of cognitive reso...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
152072be-9cdb-50fc-a6cf-26b34cf0e9c9
DOI
10.1101/2025.10.14.25338039
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Beyond Accuracy: An Efficiency- and Safety-Aware Framework for Evaluating Clinical AI with Large Language ModelsDOI 10.1101/2025.10.14.25338039
Select a neighboring publication to make it the new centre.