Back to search

Article

Screening Feedback for Language Models with Costly Verification

2026-05-06

Abstract excerpt

<title>Abstract</title> <p>Language model training and alignment rely on high-quality human feedback, yet platforms must incentivize valuable contributions while limiting harmful feedback from non-experts. We study a simple screening environment in which a platform commits to a uniform incentive policy $(\rho,R,P)$---a verification rate, a reward for submitting feedback, and a penalty imposed when verified feedba...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
a40dd2b4-24a7-5489-b4a4-aff95026620f
DOI
10.21203/rs.3.rs-8771074/v2
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Screening Feedback for Language Models with Costly VerificationDOI 10.21203/rs.3.rs-8771074/v2
Select a neighboring publication to make it the new centre.