Back to search

Article

Evaluating Large Language Models on Medical Evidence Summarization

2023-04-24

Abstract excerpt

Recent advances in large language models (LLMs) have demonstrated remarkable successes in zero- and few-shot performance on various downstream tasks, paving the way for applications in high-stakes domains. In this study, we systematically examine the capabilities and limitations of LLMs, specifically GPT-3.5 and ChatGPT, in performing zero-shot medical evidence summarization across six clinical domains. We conduct...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
920965c4-bf60-5881-94b3-a3e3f39444f8
DOI
10.1101/2023.04.22.23288967
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Evaluating Large Language Models on Medical Evidence SummarizationDOI 10.1101/2023.04.22.23288967
Select a neighboring publication to make it the new centre.