@Adidotdev They’re all struggling at the moment - Gemini & Grok more or less irrelevant - Opus theoretically a better strategist and Codex a better coder.. but both companies are playing games with their user base and pissing everyone off. China will swoop in soon with a killer open source.
@TheAmolAvasare You lot are so full of 💩 - “screenshot” - it’s literally live on your website. The 2% is a complete lie. Why wouldn’t you announce this beforehand. Of course people are going to notice. You do not realise the amount of trust you’ve just eradicated. Claude was already on thin ice
🚨 BREAKING: CLAUDE JUST GOT NERFED.
AMD’s AI director just analyzed 6,852 Claude Code sessions, 234,760 tool calls, and 17,871 thinking blocks.
Her conclusion: “Claude cannot be trusted to perform complex engineering tasks.”
Thinking depth dropped 67%. Code reads before edits fell from 6.6 to 2.0. The model started editing files it hadn’t even read.
Stop-hook violations went from zero to 10 per day.
Anthropic admitted they silently changed the default effort level from “high” to “medium” and introduced “adaptive thinking” that lets the model decide how much to reason.
No announcement. No warning.
When users shared transcripts, Anthropic’s own engineer confirmed the model was allocating ZERO thinking tokens on some turns.
The turns with zero reasoning? Those were the ones hallucinating.
AMD’s team has already switched to another provider.
But here’s what most people are missing.
This isn’t just a Claude story.
AMD had 50+ concurrent sessions running on one tool.
Their entire AI compiler workflow was built around Claude Code. One silent update broke everything.
That’s vendor lock-in. And it will keep happening.
→ Every AI company will optimize for their margins, not your workflow
→ Today’s best model is tomorrow’s second choice
→ If your workflow can’t survive a provider switch, you don’t have a workflow. You have a dependency
The fix is simple: stay multi-model.
→ Use tools like Perplexity that let you swap between Claude, GPT, Gemini in one interface
→ Learn prompt engineering that works across models, not tricks tied to one
→ Test alternatives monthly because the rankings shift fast
Laurenzo said it herself: “6 months ago, Claude stood alone. Anthropic is far from alone at the capability tier Opus previously occupied.”
Never let one vendor own your productivity.
CLAUDE OPUS 4.6 IS NERFED.
BridgeBench just proved it.
Last week Claude Opus 4.6 ranked #2 on the Hallucination benchmark with an accuracy of 83.3%.
Today Claude Opus 4.6 was retested and it fell to #10 on the leaderboard with an accuracy of only 68.3%.
A 98% increase in hallucination.
https://t.co/ttnnwBYerW just confirmed that Claude Opus 4.6 has reduced reasoning levels and is nerfed.
@openclaw@davemorin I don’t really get the opus hype. 5.4 has always been smarter, opus just had the rounded personality. If 5.4 is set up well, IMO it outperforms opus. Can always run things past opus in a native claude instance if needed.
@iamandrewz@Franzferdinan57 What are the benefits even of coding in openclaw? Seems like context is perpetually bloated, standalone codex seems to perform far better for fewer tokens
@naval Not yet bro. An agent left to its own devices will expedite slop, it needs direction. The AI powered saas is the sweet spot until the models become a lot smarter
@StephanFerraro@AlexFinn@IamEmily2050 Yeah the rumour is that they’re about to release the M5 Ultra Mac Studio with 1TB RAM, so slowing production on the previous model