Article
Evaluating Large Language Models on Medical Evidence Summarization
2023-04-24
Abstract excerpt
Recent advances in large language models (LLMs) have demonstrated remarkable successes in zero- and few-shot performance on various downstream tasks, paving the way for applications in high-stakes domains. In this study, we systematically examine the capabilities and limitations of LLMs, specifically GPT-3.5 and ChatGPT, in performing zero-shot medical evidence summarization across six clinical domains. We conduct...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 920965c4-bf60-5881-94b3-a3e3f39444f8
- DOI
- 10.1101/2023.04.22.23288967
