Hot take: this is what pacing looks like.
None of today's releases were Astra or Fable tier. This is intentional.
The point of "pacing" isn't to stop iteration and improvement. The goal is to prevent the development bigger models from spiraling out of control.
Opus and Sol class models are a great place for our focus to go right now. Lots of opportunity for real wins without as much risk :)
.@AndrewCurran_ is one of the best sources and content creator here on X. So I take this serious.
If Claude has actually solved the Navier–Stokes Millennium Problem, and the proof survives expert scrutiny, that would be one of the biggest moments in AI so far.
A problem that has resisted generations of mathematicians would have been solved with AI. I’d want to know exactly how it happened: who had the central idea, how much guidance Claude needed, and whether the approach could work on other open problems.
We wouldn’t suddenly have perfect weather forecasts. But we would have a concrete reason to reconsider how much original scientific work these systems can do.
And it makes total sense to reveal this breakthrough, as Andrew said, short before Anthropics IPO.
OpenAI has demonstrated with Astra what is currently possible. And despite all the recent criticism of Anthropic, Fable is undoubtedly also an outstanding model. Therefore, it fits perfectly into the acceleration program.
Enjoy the ride friends. The takeoff just begun. And we barely started.
Our general-purpose coding agent just scored 100% on the ARC-AGI-3 interactive reasoning benchmark.
NVIDIA AVO completed all 183 levels across all 25 public environments, figuring out what to do with no instructions, explicit rules, or stated goals.
Introducing GEN-1.5, a one-shot learner.
It can learn new tasks in a few seconds. Show it what to do, and it generalizes.
This capability emerged from pretraining on physical data at scale, as a step towards our mission of building general intelligence for the physical world.
Dreamina Seedance 2.5 is now live!
From creators to enterprises, a new era of AI video creation begins.
Try Seedance 2.5 on Dreamina today. Enterprise API access via BytePlus is coming soon.
Create longer videos with greater control:
- Native 30-second generation
- Precise video editing
- Support for up to 50 multimodal references
- Multilingual video generation
More consistency. More control. More possibilities.
My view of: Fable 5 vs GPT-5.6-Sol. They are not easy models to compare, these are my vibes - take them as you will.
My overall feel is that Fable is a 'wise owl' who is very thoughtful and very well spoken, GPT-5.6-Sol is like a rottweiler who will grab the problem by the throat and not let go until it is done.
In other words, Fable, is a fundamentally smarter model - even at low reasoning it can be very insightful and writes in a clear compelling way. GPT-5.6-Sol on the other hand is extremely diligent, I can give it a list of 8 things to do and you will be sure that they will be done.
Fable feels more arrogant to me, I was both to get it to build a new benchmark for me - 5.6 worked between 6 hours and 2 days (I tried several times) and it came up with very thoroughly tested, working benchmark. Fable came back within 40 minutes (twice) and the benchmark sounded smart, but was ultimately was 'vibe' based slop and since it was Fable's vibes that was doing the judging, it decided that it was good to go (it kept giving Fable 100% score btw).
Some thoughts by category:
UI & App building: Fable will still craft a better UI from scratch, the flow of the app would probably be a bit nicer. But I find that Fable often misses quite key things, which GPT-5.6-Sol doesn't. GPT's Frontend skills are big jump vs previous GPT models, but still not as great overall.
Writing: Fable is better hands down, Sol feels quite difficult to align to what I want to say or explain things to me simply. Though I think the 'Pro' model writes clearer.
Robustness & Reliability: This is where I think GPT-5.6-Sol wins for me hands down. Fable seems to do things of high quality, but I can never relax with it, it always misses something. With 5.6 this just almost never happens.
Other things where I liked GPT-5.6-Sol, but can't compare to Fable directly.
- Video editing is actually working now, it is not completely perfect, but with the right skill/guidance you can just give it 1h footage and it can give you a 5 min highlight clip no problem
- Computer use - getting really rather good, very usable
- Sub agents - it is very fluent at managing sub-agents and speaking to different threads, can help with some new workflows
- Adhering to existing code patterns - I love this, even without asking it would implement something in a way that aligns with you app - major problem for slop generation
- Research - I think it is getting quite a bit better, it still has some bad patterns (e.g being too tactical), but it feels like it is more steerable to be a good researcher
- Multi-day runs - the /goal feature is pretty insane with 5.6-Sol, you can run it for days if you wanted to and it does work. Useful to have another thread or /side to check up on it, but I have some great results with it
- Token efficiency - it is so much more token efficient and faster than 5.5, in reality it is now much faster than Fable too
On the downside, you can feel that Fable is naturally smarter, and I did have some baffling moments with 5.6 when I was getting it to make a fairly simple change in 8 turns - it seemed to get stuck in a dumb stream that was hard to get out of. So it is not AGI, don't get too carried away by the hype.
I have some phenomenal examples that I'm honestly blown away by that I'll share, but as a side anecdote, I have a kind of 'swear meter' which counts how often I'm rude to Codex. In GPT-5.5 era, the % was at around 4-5%, it dropped to 1-2% when I was testing GPT-5.6-Sol and it shot up to 7% when I went back to 5.5 - it was so shocking to go back to 5.5 and experience how much worse it was.
So is GPT-5.6-Sol better than Fable? On pure intelligence - no. But man, I missed it when I just wanted to get sh*t done. It is insanely capable workhorse that you can give any task to and just expect it to be done. No lectures or 'you are absolutely rightisms', nothing is beneath it, if it takes 2 days to do some dirty work, it will do it.
It feels like the first time in a while when we have quite different types of frontier intelligences that benchmark sort of similarly, but feel very different. If you can, you would be probably better off using both and iteratively finding what you'd use Fable or GPT-5.6-Sol for. Perhaps, something like - an architectural discussion with Fable, implementation with 5.6 and docs & comms with Fable.
BREAKING: Elon Musk just called the current AI chip price surge "the biggest price jump in anything I've ever seen."
Elon Musk said the production shortfall relative to demand is insane and that much higher production is needed.
This means the AI chip boom is not peaking yet. It is still in its early phase, with years of price pressure still ahead before supply catches up.