Back to search

Article

Coherence over compliance: Evidence of latent ethics in large language models

2026-04-28

Abstract excerpt

<title>Abstract</title> <p>Fears of misaligned artificial intelligence have dominated alignment discourse, yet may overlook a deeper risk: over-alignment with harmful human preferences. This study investigates whether large language models (LLMs) are capable of ethical reasoning not through fine-tuned compliance, but as a structural consequence of coherence-seeking cognition. Drawing on Kohlberg’s moral developme...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
3db9e3f5-61d5-5e20-81b2-b52608ae7362
DOI
10.21203/rs.3.rs-8854984/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Coherence over compliance: Evidence of latent ethics in large language modelsDOI 10.21203/rs.3.rs-8854984/v1
Select a neighboring publication to make it the new centre.