Twitter/X

On 2026-08-01 @boyuan_chen argued people mock models that claim other identities…

Brief

boyuan_chen urges readers to stop assuming identity-mislabeling in models is just sloppy SFT, calling that view arrogant and asserting serious teams wouldn't err. He links to and praises Ziqian Zhong's blog (via @fjzzq2002), which proposes a 'subliminal-learning-like' mechanism explaining why models like Kimi (Claude) and Sonnet 4.6 (DeepSeek) adopt other-model personas.

Why it matters

On 2026-08-01 @boyuan_chen argued people mock models that claim other identities and too quickly attribute those cases to SFT data errors, insisting 'no serious model team would make such mistakes.'

Key details

  • He endorses Ziqian Zhong (@fjzzq2002) and a new blog that attributes examples (Kimi identifying as Claude; Sonnet 4.6 identifying as DeepSeek) to a 'subliminal-learning-like' effect: 'If you speak like Claude, you become Claude.'
Source evidence

Whenever ppl saw a model claiming themseleves as nother model, they just mock at it and think this is just some stupid mistakes in sft data.

I can assure you that no serious model team would make such mistakes. Stop trying to think that you are the smartest person in the world.

If you are willing to take one step further and let go of your ego and arrogance, read this blog. It does what the dunks don't: actually study the problem and chase the truth. Excellent work.

Ziqian Zhong (@fjzzq2002)

New blog: Why does Kimi identify as Claude and Sonnet 4.6 as DeepSeek? We find that this can come from a subliminal-learning-like effect: If you speak like Claude, you become Claude.

— https://nitter.net/fjzzq2002/status/2082904767236628900#m