ICYMI we've cut the price of Sonnet 5.5 cache reads in half.
In the Claude Platform, it's now $0.10 per million tokens (input is $2, output is $10).
This means Sonnet 5.5 now runs ~20% cheaper on most agentic work.
(API only, no change to Claude Code usage limits.)
Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we’ve ever released.
On average, it costs around 75% less to run than Claude Haiku 4.5.
🚨 fable 5.5 Leaks: Beats Opus 5.5
> Fable 5.5 is already be secretly rolling out
> Fable 5.5 appeared in Anthropic’s docs
> Some users are being silently routed to Fable 5.5 while the UI still says Fable 5.1
> The first demos are showing a major jump in capability
> Early tests it is outperforming Opus 5.5 and GPT-6 Astra on some tasks
> Anthropic may be testing the model with a limited group before a wider release
> A public launch could potentially happen as soon as possible
Are we actually days away from Fable 5.5?
Claude Opus 5.5 just took a big drop on NerfBench.
Yesterday it was scoring above launch. Today it's at 94.2%.
GPT 6 Astra: 98.0%
Sonnet 5.5: 100.9%
GPT 6.1 Sol: 106.7%
94.2% is still inside normal variance, so we can't call it a nerf yet.
But we're watching Opus 5.5 very closely.
Big news: Gemini 4 Argon (High) by @GoogleDeepMind just landed #1 in Text Arena with 1525 pts, and #8 in Code Arena: WebDev with 1679 pts!
This release has reshaped the Text Arena Pareto frontier with a blended $8/MToken! Gemini 4 Argon (High) is now the most cost efficient model, see its placement on Pareto frontier below.
In the Text Arena, Gemini 4 Argon (High) ranks #1 in Coding, Hard Prompts, Instruction Following, Longer Query, and Creative Writing. It also leads every occupational domain evaluated, with additional #1 spots in English, Non-English, Chinese, and Russian.
This model is +20 points above the #2 ranked Claude Opus 4.6 (High), and a huge leap from Google’s previous release, Gemini 3.8 Flash (High) at #11!
In Code Arena: WebDev, Gemini 4 Argon (High) gained +96 points from Gemini 3.8 Flash (High), and went from #29 to #8.
Congrats to the @GoogleDeepMind team on this impressive frontier release!
Introducing Gemini 4 Argon – our new frontier model.
It’s built for complex workflows across coding, enterprise knowledge work, and cybersecurity defense – rolling out today to a set of trusted testers through our Fairwind Program.