Back to search

Article

The Moral Turing Test: Evaluating Human-LLM Alignment in Moral Decision-Making

2024-10-01

Abstract excerpt

<p>As large language models (LLMs) become increasingly integrated into society, their alignment with human morals is crucial. To better understand this alignment, we created a large corpus of human and LLM-generated responses to various moral scenarios. We found a misalignment between human and LLM moral assessments; although both LLMs and humans tended to reject morally complex utilitarian dilemmas, LLMs were mor...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
ecd36797-da26-5769-9754-679e6e4c957d
DOI
10.31234/osf.io/ct6rx
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
The Moral Turing Test: Evaluating Human-LLM Alignment in Moral Decision-MakingDOI 10.31234/osf.io/ct6rx
Select a neighboring publication to make it the new centre.