Article
WITHDRAWN: Alignment-Induced Flattening in Large Language Models: A 33-Item Reading Comprehension Measure of Atypical Perspective-Taking Across Five Model Families
2026-02-17
Abstract excerpt
<title>Abstract</title> <p> Large Language Models (LLMs) fine-tuned with Reinforcement Learning from Human Feedback (RLHF) are known to exhibit pro-social biases in their outputs <sup>4,5</sup> . We hypothesise that this alignment process produces a measurable artefact in narrative interpretation: <bold>alignment-induced flattening</bold> , whereby models systematically favour empathetic, romantic, or morall...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- e41eabc8-d0f5-5d5d-b879-d6f641cce4bb
- DOI
- 10.21203/rs.3.rs-8879713/v1
