Back to search

Article

Collaborative intelligence in AI: Evaluating the performance of a council of AIs on the USMLE

2025-02-20

Abstract excerpt

The variability in responses generated by Large Language Models (LLMs) like OpenAI’s GPT-4 poses challenges in ensuring consistent accuracy on medical knowledge assessments, such as the United States Medical Licensing Exam (USMLE). This study introduces a novel multi-agent framework—referred to as a "Council of AIs"—to enhance LLM performance through collaborative decision-making. The Council consists of multiple...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
53820eea-1d0b-58dd-ae80-608d4edbabd3
DOI
10.1101/2025.02.17.25322388
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Collaborative intelligence in AI: Evaluating the performance of a council of AIs on the USMLEDOI 10.1101/2025.02.17.25322388
Select a neighboring publication to make it the new centre.