Back to search

Article

“Mirror” Language AI Models of Depression are Criterion-Contaminated

2025-08-07

Abstract excerpt

<p>A growing number of studies show near-perfect LLM language-based prediction of depression assessment scores (up to R2 of .70). However, many are developed from language responses directly to the depression assessments. These “Mirror models” suffer from “criterion contamination,” which arises when a predicted score depends in part on the predictors themselves. This causes artificial effect size inflation which r...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
a92ab810-3403-5de2-b2f1-7640659b67af
DOI
10.31234/osf.io/fe2z9_v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
“Mirror” Language AI Models of Depression are Criterion-ContaminatedDOI 10.31234/osf.io/fe2z9_v1
Select a neighboring publication to make it the new centre.