Back to search

Article

Validating local open-weight LLMs against expert raters for graded deductive coding of open-ended responses in Russian

2026-08-07

Abstract excerpt

<p>Large language models (LLMs) are increasingly used to code open-ended survey responses, but their validity against expert annotation remains uncertain, especially outside English. We evaluated four open-weight LLMs for deductive coding of Russian responses from a psychological human-robot interaction experiment (N = 33): DeepSeek-R1-32B, Qwen3-32B, Gemma3-27B, and Mistral-Small-3.2-24B. Ten psychology-trained c...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
22db7a3b-392c-5950-8c47-7378abc2e96e
DOI
10.31234/osf.io/8psd9_v2
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Validating local open-weight LLMs against expert raters for graded deductive coding of open-ended responses in RussianDOI 10.31234/osf.io/8psd9_v2
Select a neighboring publication to make it the new centre.