Back to search

Article

chatRater: Validating LLM-Generated Psycholinguistic Norms with Probability-Weighted Scoring

2026-05-16

Abstract excerpt

<p>Large language models (LLMs) offer a scalable alternative to labour-intensive human normingstudies in psycholinguistic research. This article introduces chatRater, an R package for ratingtext, image, and audio stimuli via multiple LLM providers, and demonstrates its use with a casestudy validating GPT-4o ratings of English idioms against human norms. We extend theprobability-weighted scoring method of Brysbaert...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
c5954f5a-ec6c-52ef-a1e9-a879fd1677e2
DOI
10.31234/osf.io/mje6w_v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
chatRater: Validating LLM-Generated Psycholinguistic Norms with Probability-Weighted ScoringDOI 10.31234/osf.io/mje6w_v1
Select a neighboring publication to make it the new centre.