Back to search

Article

WITHDRAWN: Alignment-Induced Flattening in Large Language Models: A 33-Item Reading Comprehension Measure of Atypical Perspective-Taking Across Five Model Families

2026-02-17

Abstract excerpt

<title>Abstract</title> <p> Large Language Models (LLMs) fine-tuned with Reinforcement Learning from Human Feedback (RLHF) are known to exhibit pro-social biases in their outputs <sup>4,5</sup> . We hypothesise that this alignment process produces a measurable artefact in narrative interpretation: <bold>alignment-induced flattening</bold> , whereby models systematically favour empathetic, romantic, or morall...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
e41eabc8-d0f5-5d5d-b879-d6f641cce4bb
DOI
10.21203/rs.3.rs-8879713/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
WITHDRAWN: Alignment-Induced Flattening in Large Language Models: A 33-Item Reading Comprehension Measure of Atypical Perspective-Taking Across Five Model FamiliesDOI 10.21203/rs.3.rs-8879713/v1
Select a neighboring publication to make it the new centre.