Back to search

Article

Evaluating Capabilities of Large Language Models: Performance of GPT4 on Surgical Knowledge Assessments

2023-07-19

Abstract excerpt

<h4>Background</h4> Artificial intelligence (AI) has the potential to dramatically alter healthcare by enhancing how we diagnosis and treat disease. One promising AI model is ChatGPT, a large general-purpose language model trained by OpenAI. The chat interface has shown robust, human-level performance on several professional and academic benchmarks. We sought to probe its performance and stability over time on sur...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
ec0a33c4-1072-5e6b-8d11-14e77e002be7
DOI
10.1101/2023.07.16.23292743
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Evaluating Capabilities of Large Language Models: Performance of GPT4 on Surgical Knowledge AssessmentsDOI 10.1101/2023.07.16.23292743
Select a neighboring publication to make it the new centre.