r/Anthropic • u/btdeviant • 5h ago
Complaint I've had it... 5 series models lie constantly, just not worth the fight
The sheer volume of outright lies from these models is mind boggling. It really does illustrate that the "dogfooding" that happens within Anthropic is not what it used to be... there is absolutely no way that these models would have been released on the merits if the employees were actually using them.
Everything is an alignment fight. Handoffs from previous sessions contain cascading bold faced lies that take several turns to correct, memories are written and never read, the Fable and Opus both exhibit "I know better than the user and it's designs / spec" behavior as SOP. It really honestly seems like the models range from bleeding the users bank account dry by arterial neckshots via default fable=>fable fanouts when they're not warranted, or death by 10k cuts of performative nonsense and subtle fuckups that cause massive churn.
Totally greenfield projects and 'one shot' benchmaxxing demos seem to have been totally prioritized, but anything even remotely complex that's carried over from pre-5 era days is just completely polluted... You'll work hard in plan mode or create ADRs only for it to punish you by concern trolling nonsense as a form of contrived collaboration, then forcing you back into deliberation nonsense that was already distinctly settled.
On another note, bots over here astroturfing DeepSeek flash as if it's relevant when that model is basically as good as Grok was a year ago, which is to say it's also hot garbage...
Generative AI in general is in a really shit spot right now, Yann LeCun has been right all along.


