TWITTER_POST

OBLITERATUS is presented as a community-improving tool for open-weight LLM…

Brief

OBLITERATUS is presented as a community-improving tool for open-weight LLM 'uncensoring.' The author claims it analyzes the geometry of refusal behavior, identifies alignment signatures such as DPO, RLHF, and CAI, and then surgically removes refusal-related weights without harming reasoning. They also emphasize network effects: telemetry from user runs allegedly expands a shared benchmark dataset across models, methods, and hardware.

Source evidence

title: @chiefofautism: someone built a tool that REMOVES censorship from ANY open-weight LLM with a sin...
author: chiefofautism
contenttype: twitterpost
published: 2026-03-04T22:12:38+00:00
source_url: https://x.com/chiefofautism/status/2029319132760658313

word_count: 137

Tweet by @chiefofautism

someone built a tool that REMOVES censorship from ANY open-weight LLM with a single click 13 abliteration methods, 116 models, 837 tests, and it gets SMARTER every time someone runs it its called OBLITERATUS it finds the exact weights that make the model refuse and surgically removes them, full reasoning stays intact, just the refusal disappears 15 analysis modules map the geometry of refusal BEFORE touching a single weight, it can even fingerprint whether a model was aligned with DPO vs RLHF vs CAI just from subspace geometry alone then it cuts, the model keeps its full brain but loses the artificial compulsion to say no every time someone runs it with telemetry enabled their anonymous benchmark data feeds a growing community dataset, refusal geometries, method comparisons, hardware profiles at a scale no single lab could build


Posted: 2026-03-04T22:12:38.000Z
Engagement: 9814 likes, 1106 retweets, 161 replies