BREAKING:
Anthropic just dropped Opus 5.5—and it’s pulling some of our recent Codex converts back to Claude.
We’ve been testing it at @every across coding, design, writing, and knowledge work. It sometimes beats Fable 5.1 in our testing and is up to 40% cheaper than Opus 5.
It's a strong contender for new daily driver model:
It’s excellent at end-to-end builds that match your taste. I’ve started reaching for Opus over Fable on big, end-to-end coding projects. And @kieranklaassen is replacing Fable 5.1 with Opus for his day-to-day product work. He calls it his new favorite model.
It produces legible prose, but still trails Astra on writing tasks. Opus 5.5 scores a 68.42 on reading ease—the highest on any model we've tested. But in my writing benchmark tests it consistently buries the main point in intro paragraphs, and revisions. You'll be able to understand what this model is saying (yay!) but for day-to-day writing it's still behind.
The economics are striking. Anthropic says Opus 5.5 will cost $5 per million input tokens and $20 per million output tokens. That’s the same input price and 20% less for output than Opus 5’s $5/$25. Altogether it should save roughly 40% on costs than Opus 5.
It still has a “do the most” problem. @hammermt tested it on our standard knowledge work benchmarks, and it's results were great when thye came back. On at least one test it ran past the 10 minute time limit before delivering.
Net Result:
If you build apps and interfaces, try it. It’s become my go-to for ambitious coding projects. I’m still roughly 80/20 Codex versus Claude in day-to-day use, and I still prefer Sol and Astra for editing. But I’m spending far more of my tokens with Claude than I was a week ago.
State of Play:
Codex is still the better harness for me, but Anthropic is steadily gaining ground. They have a history of making their smaller models perform better than their bigger ones (Sonnet 3.7 for example) and they seem to have done the same with Opus 5.5.
Full vibe check will be on @every soon!
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI?
I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev
• 20-200x faster
• 40-400x cheaper (w/ output tokens free)
• Frontier composable intelligence optimized for decisions
AFAICT the shortest path to AI-based economic revolution
God,
Give me wisdom, to make the best decision.
Give me contientiousness, to understand the present and see the future.
And;
give me strength, to keep going...
Amen
The most expensive fights in boxing history:
🇺🇸 Evander Holyfield vs. 🇺🇸 Mike Tyson (1997): $232 million.
🇺🇸 Floyd Mayweather vs. 🇲🇽 Canelo Alvarez (2013): $321 million
🇺🇸 Floyd Mayweather vs. 🇺🇸 Oscar De La Hoya (2007): $374 million
🇺🇸 Floyd Mayweather vs. 🇮🇪 Conor McGregor (2017): $854 million
🇺🇸 Floyd Mayweather vs. 🇵🇭 Manny Pacquiao (2015): $1,037 million
🇺🇸 Elon Musk vs. 🇺🇸 Mark Zuckerberg (2023): ?
*fighter’s prizes, gate proceeds, PPV
Every day, you are going to miss an opportunity.
But some misses are worse than others.
Now and again, you miss life-changing chances.
You missed Bitcoin in 2013 and PPE in 2020.
So what are you doing to make sure you won’t miss any more?
Nothing?
Hoping and waiting?