I’m convinced that we’re building more and more dependencies on LLM’s to the point at which the smallest degradation will feel like the end of the world.
It’s also insane how a small weight somewhere has turned the model into a lazy, degraded co-worker instead of the A-player it was just a short while ago.
the model whisperers at the big labs have way more important jobs than the model makers realize.
small scale issues like the laziness described by Kim here quickly snowball out of control and the butterfly effect is in full swing.
Nobody wants to pay for laziness, even if it’s just perceived laziness. it’s simple human psychology.
Seeing tons of people sharing the same sentiment across Claude and ChatGPT latest model releases.
it’s time for the labs to start treating model behaviors like the single most important vector.
Chubby♨️ (@kimmonismus)
I'm going to cancel Claude. It's just so bad, I can't believe it.
It's just lazy. The most recent example: I have Claude check my inbox for important emails, summarize them, work with them, and send out replies if necessary. I caught Claude again simply not reading the email thread to the end and just ignoring the latest emails.
When I asked him about it, Opus 5 just said: "Valid point. I didn't read it."
I mean, seriously. What the heck? You have to babysit it every time.
— https://nitter.net/kimmonismus/status/2084727476761083926#m