Article
Rx-LLM: a benchmarking suite to evaluate safe large language model performance for medication-related tasks
2025-12-02
Abstract excerpt
<h4>Background:</h4> For large language models (LLMs) to reach their potential as information technology tools that make medication use safer, clinically relevant benchmarks capable of automated grading and designed specifically to measure the performance of LLMs for medication tasks are required. The purpose of this study was to design a suite of benchmarking tests reflective of Comprehensive Medication Manageme...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- f1a36e9c-ca9a-5d3d-9c08-d49101be2cbd
- DOI
- 10.64898/2025.12.01.25341004
