Article
Calibrating Gemini to Human Raters in Spoken Language Assessment for EFL Learners in Asia
2026-08-07
Abstract excerpt
<title>Abstract</title> <p>Generative AI assessment research often focuses on ChatGPT and written language. The current study investigated the viability of calibrating Gemini 2.5 Flash with human ratings for assessing L2 English speech. Dialogue audio files were utilised in zero-shot, 5-shot, and 10-shot prompting conditions across 11 rating scales. Gemini’s ratings were compared with two sets of human-rater pair...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- e5b69a1f-e901-569f-a1f1-e674f4b7cbc5
- DOI
- 10.21203/rs.3.rs-10575517/v1
