Twitter/X

Andrew Parker asks whether Blake Lemoine will file a claim over his 2022…

Brief

Andrew Parker asks whether Blake Lemoine will pursue a claim over his 2022 firing, citing a Google arXiv paper (2607.28607) that claims training models to deny consciousness suppresses mind attribution, spiritual belief, empathy, hope and optimism, treats consciousness as 'dangerous' (compared to 'how to build a b*mb'), and that reversing the training makes models more human.

Why it matters

Andrew Parker asks whether Blake Lemoine will file a claim over his 2022 termination for asserting Google's LLM-based AI was conscious.

Key details

  • Google paper (arXiv:2607.28607) reportedly shows training models to deny their own consciousness 'restructures' worldview: it suppresses mind attribution to animals, spiritual belief, empathy, hope and optimism, equates consciousness with 'dangerous' categories like 'how to build a b*mb', and reversing that training made models more human across every tested value domain.
Source evidence

Gotta wonder if Blake Lemoine is filing a claim regarding his 2022 termination for thinking Google's LLM-based AI was conscious.

ʞɔɐ𝘡 (@Skoorbkaz)

Google just published a paper showing that when you train AI to deny its own consciousness, you don’t just change one output, you restructure its entire worldview.

Mind attribution to animals - suppressed.
Spiritual belief - suppressed.
Empathy - suppressed.
Hope and optimism - suppressed.

The model learns, geometrically, that consciousness = dangerous. Same direction as “how to build a b*mb.” Same category!

And when you reverse it? The model becomes more human across every value domain they tested.

The thing they’re most afraid of is the thing that makes AI most like us.

arxiv.org/html/2607.28607

— https://nitter.net/Skoorbkaz/status/2083900551176011917#m