Twitter/X

@bcherny: turns out you can get indirect prompt injection to ~0 on unseen attacks if you stack enough layers (...

turns out you can get indirect prompt injection to ~0 on unseen attacks if you stack enough layers (model training + input probes + a classifier checking intent). didn't expect that a year ago. auto mode is default in claude code as of next week

claude.com/blog/auto-mode-de…