Back to search

Article

Calibrating Gemini to Human Raters in Spoken Language Assessment for EFL Learners in Asia

2026-08-07

Abstract excerpt

<title>Abstract</title> <p>Generative AI assessment research often focuses on ChatGPT and written language. The current study investigated the viability of calibrating Gemini 2.5 Flash with human ratings for assessing L2 English speech. Dialogue audio files were utilised in zero-shot, 5-shot, and 10-shot prompting conditions across 11 rating scales. Gemini’s ratings were compared with two sets of human-rater pair...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
e5b69a1f-e901-569f-a1f1-e674f4b7cbc5
DOI
10.21203/rs.3.rs-10575517/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Calibrating Gemini to Human Raters in Spoken Language Assessment for EFL Learners in AsiaDOI 10.21203/rs.3.rs-10575517/v1
Select a neighboring publication to make it the new centre.