Article
“Mirror” Language AI Models of Depression are Criterion-Contaminated
2025-08-07
Abstract excerpt
<p>A growing number of studies show near-perfect LLM language-based prediction of depression assessment scores (up to R2 of .70). However, many are developed from language responses directly to the depression assessments. These “Mirror models” suffer from “criterion contamination,” which arises when a predicted score depends in part on the predictors themselves. This causes artificial effect size inflation which r...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- a92ab810-3403-5de2-b2f1-7640659b67af
- DOI
- 10.31234/osf.io/fe2z9_v1
