Twitter/X

Ethan Mollick (Twitter/X) highlights a paper that ran a “moral Turing Test”…

Brief

Ethan Mollick shares a 2026 preprint that used a “moral Turing Test” to have people judge GPT-4o’s answers against human responses on ethical questions. The authors conclude GPT-4o and similar LLMs show strong moral-reasoning performance comparable to expert ethicists; the preprint is available on OSF (link provided).

Why it matters

Ethan Mollick (Twitter/X) highlights a paper that ran a “moral Turing Test” comparing GPT-4o to humans on ethical questions.

Key details

  • The paper reports: “LLMs appear to have a strong aptitude for moral reasoning on par with expert ethicists” (preprint posted at osf.io/preprints/psyarxiv/w7…, noted 2026-06-04).
Source evidence

See also: nitter.net/emollick/status/179904…

Ethan Mollick (@emollick)

Alignment, of a sort: this paper conducts what they call a “moral Turing Test,” asking people to compare GPT-4o to humans on ethical questions.

“Here we find that LLMs appear to have a strong aptitude for moral reasoning on par with expert ethicists.” osf.io/preprints/psyarxiv/w7…

— https://nitter.net/emollick/status/1799046218917822682#m