Back to search

Article

Evaluating reasoning LLMs’ potential to perpetuate racial and gender disease stereotypes in healthcare

2025-08-07

Abstract excerpt

This evaluation of 36,000 clinical vignettes found that next-generation reasoning large language models, o3-mini and DeepSeek-R1, frequently perpetuate racial and gender stereotypes for common medical conditions, indicating that advancements in reasoning do not inherently improve representational fairness.

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
9d294710-db98-5033-be7e-654876a4ea3f
DOI
10.1101/2025.08.05.25333007
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Evaluating reasoning LLMs’ potential to perpetuate racial and gender disease stereotypes in healthcareDOI 10.1101/2025.08.05.25333007
Select a neighboring publication to make it the new centre.