Article
Validating local open-weight LLMs against expert raters for graded deductive coding of open-ended responses in Russian
2026-08-07
Abstract excerpt
<p>Large language models (LLMs) are increasingly used to code open-ended survey responses, but their validity against expert annotation remains uncertain, especially outside English. We evaluated four open-weight LLMs for deductive coding of Russian responses from a psychological human-robot interaction experiment (N = 33): DeepSeek-R1-32B, Qwen3-32B, Gemma3-27B, and Mistral-Small-3.2-24B. Ten psychology-trained c...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 22db7a3b-392c-5950-8c47-7378abc2e96e
- DOI
- 10.31234/osf.io/8psd9_v2
