Back to search

Article

Rx-LLM: a benchmarking suite to evaluate safe large language model performance for medication-related tasks

2025-12-02

Abstract excerpt

<h4>Background:</h4> For large language models (LLMs) to reach their potential as information technology tools that make medication use safer, clinically relevant benchmarks capable of automated grading and designed specifically to measure the performance of LLMs for medication tasks are required. The purpose of this study was to design a suite of benchmarking tests reflective of Comprehensive Medication Manageme...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
f1a36e9c-ca9a-5d3d-9c08-d49101be2cbd
DOI
10.64898/2025.12.01.25341004
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Rx-LLM: a benchmarking suite to evaluate safe large language model performance for medication-related tasksDOI 10.64898/2025.12.01.25341004
Select a neighboring publication to make it the new centre.