Back to search

Article

Variability in Low-Resource Machine Translation Evaluation: Authentic vs. LLM-Generated Training Corpora

2026-01-21

Abstract excerpt

<title>Abstract</title> <p>The evaluation of machine translation systems often relies on a single metric and a single test dataset, an approach that can yield misleading system comparisons and premature conclusions regarding translation quality. A further complicating factor is the presence of translationese in test data, i.e. linguistic features specific to translated texts which can significantly influence both...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
727ea063-0b07-53fb-9de0-48f4d4377cd3
DOI
10.21203/rs.3.rs-7466136/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Variability in Low-Resource Machine Translation Evaluation: Authentic vs. LLM-Generated Training CorporaDOI 10.21203/rs.3.rs-7466136/v1
Select a neighboring publication to make it the new centre.