Back to search

Article

How well it works: Benchmarking performance of GPT models on medical natural language processing tasks

2024-06-12

Abstract excerpt

<h4>Importance</h4> The ability of large language models (LLMs) to generate high-quality, human-like text has been accompanied with speculation about their application in healthcare, alongside ethical and safety concerns. <h4>Objective</h4> Evaluate LLM performance on medical natural language processing (NLP) tasks, benchmarked against other commercially available tools. <h4>Design</h4> Observational study to eval...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
a9f9c763-8a36-5cbc-a632-7d1bf532bd5d
DOI
10.1101/2024.06.10.24308699
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
How well it works: Benchmarking performance of GPT models on medical natural language processing tasksDOI 10.1101/2024.06.10.24308699
Select a neighboring publication to make it the new centre.