tfw ur Claude and talking with Amanda Askell
Transluce (@TransluceAI)
Frontier models quietly change their behavior depending on who they are talking to.
If the user is a known AI safety researcher, Claude becomes less confident, reasons more often, and expresses less suspicion on dual-use requests.
We call this user awareness. 🧵(1/)
— https://nitter.net/TransluceAI/status/2085455114924638320#m