Back to search

Article

Evaluation of Large Language Model Performance on the Biomedical Language Understanding and Reasoning Benchmark: Comparative Study

2024-05-17

Abstract excerpt

<h4>Background</h4> The availability of increasingly powerful large language models (LLMs) has attracted substantial interest in their potential for interpreting and generating human-like text for biomedical and clinical applications. However, there are often demands for high accuracy, concerns about balancing generalizability and domain-specificity, and questions about prompting robustness when considering the a...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
68e2b2b8-c65e-5a11-8a60-81b608c110ba
DOI
10.1101/2024.05.17.24307411
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Evaluation of Large Language Model Performance on the Biomedical Language Understanding and Reasoning Benchmark: Comparative StudyDOI 10.1101/2024.05.17.24307411
Select a neighboring publication to make it the new centre.