Article
Screening Feedback for Language Models with Costly Verification
2026-05-06
Abstract excerpt
<title>Abstract</title> <p>Language model training and alignment rely on high-quality human feedback, yet platforms must incentivize valuable contributions while limiting harmful feedback from non-experts. We study a simple screening environment in which a platform commits to a uniform incentive policy $(\rho,R,P)$---a verification rate, a reward for submitting feedback, and a penalty imposed when verified feedba...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- a40dd2b4-24a7-5489-b4a4-aff95026620f
- DOI
- 10.21203/rs.3.rs-8771074/v2
