Article
Evaluating Large Language Models for Psychometric Simulation Studies in R: Integrating Best Practices in Simulation and Prompt Design
2026-04-08
Abstract excerpt
<p>While large language model (LLM) capabilities have been widely investigated across various domains, limited research has examined their capability to generate end-to-end R code for conducting simulation studies in psychometrics. To address this gap, we evaluated four commonly used reasoning models: ChatGPT-5.1 (Thinking), Claude 4.5 Opus, DeepSeek-V3.2, and Gemini 3 Pro, in generating simulation code correspond...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- db68a766-73d2-51f8-8732-5179c97c7b3b
- DOI
- 10.31234/osf.io/9tfej_v1
