Many people outside our AI community have absolutely no idea what's happening right now. I keep thinking about Demi Hassabi's words that we're entering the golden age of science. This is proof of his prediction.
What will this affect? Everything. Materials research, energy production, drug discovery.
Everything.
I don't want to get too caught up in the hype. But the fact that we could seriously cure every form of cancer, if not stop aging, at least double the normal rate, gain access to new energy sources, and so on, excites me.
The signal is no longer merely that AI can sometimes produce an ingenious proof. The signal is that discovery itself may be becoming a scalable computational process.
The ten results are not the deepest event here. They are evidence of the event.
The event is the arrival of a new engine of discovery.
(comment via GPT-5.6 Pro)
okay so. when a mathematician proves a random conjecture you are supposed to politely clap. this is because it would be rude and embarrass the mathematician to point out that most conjectures are not that important or meaningful
when an LLM proves a random conjecture mathematicians are caught in a funny position where they want to say "these are not very important or meaningful conjectures and having magic proofs or counterexamples come out of nowhere is not particularly interesting or useful to anybody, except possibly the person who proposed the conjecture and some of their colleagues" but saying that puts them in an awkward position - why didn't they say that when humans were doing it? was nearly all mathematics actually pointless this whole time and mathematicians were just politely avoiding ever pointing this out so they could all still have jobs?
if the best defense you can come up with of mathematics as a human endeavor is, effectively, "it's really fun for us to come up with these baroque puzzles and then solve them, don't screw this up for us," you are completely abdicating your responsibility as a public endeavor that taxpayers fund. i have occasionally argued that pure mathematics lost the plot decades ago and gave up on having any kind of meaningful relationship to the rest of the human quest to understand reality. this was a slow-burning tragedy, and mathematicians have been able to hide this state of affairs from everybody else because nobody else knows enough mathematics to critique what mathematicians are doing
i saw a famous professor point out a few years ago that the machine learning that underpins AI is powered by 18th or maybe 19th century mathematics, you don't need much more than multivariable calculus, linear algebra, a bit of probability. he said this as if it was a dunk on ML and AI; imo it's a much bigger dunk on mathematics. you're telling me the most important scientific revolution of our lifetimes doesn't involve any of the math that mathematicians have been working on over the last century or two? you see how that's worse, right?
(tbc i'm not saying that stuff is all useless, you need a bunch of stuff from the first half of the 20th century to do physics, e.g. functional analysis, lie theory, representation theory all feed into the standard model. after 1950 or so things get more dicey)
if mathematicians still want to have jobs they are going to have to do a better job than they have for the last few decades convincing the public they do something that actually matters, and figuring out what that actually is or should be. this is an opportunity for the field to do some very serious soul-searching, as serious or more serious than the foundational crisis from a century ago. this kind of whining is absolutely not going to cut it
WOWII Conjecture 91 is false. This graph theory problem was open for ~22 years.
I cracked it by asking ChatGPT 5.6 Pro to find a counterexample to an open conjecture of its choosing and then I went to have a nap.
I do NOT care for graph theory nor had I heard of this NERD conjecture until ChatGPT told me it found a solution.
ChatGPT conversation where this was found:
https://t.co/4AnMhde026
there are people who have load-bearing beliefs about human intelligence generally and their intelligence specifically that increasing AI capabilities render increasingly obviously fake, and they are not going to have a good time. the truth is that the vast majority of what we call thinking is also just regurgitating or recombining thoughts we've seen already, what frontier models can already do is simply not that different from what we do nearly all the time. your highly original thoughts that demonstrate the depth and breadth of your soul are mostly warmed-over reheated leftovers of leftovers of thoughts that were first articulated centuries if not millennia ago. this is completely fine and ordinary and not a problem, unless you were basing your ego defenses on this very popular but largely fake misunderstanding of what thoughts are and where they come from. you cannot own a thought. they were never yours to accrete your sense of self around. they are farts that gently pass through the butthole of your mind, full of sound and fury, signifying nothing
'Execution is becoming abundant, fast, and scalable. The complement to execution is verification: the capacity to know whether what was executed is what was intended.'
I've always been skeptical of the rosy picture painted by some people of what post-AGI employment will look like. In fact, I've said before I think it's going to hit human employment like a meteor.
However, if what's going on with math right now is any indication, we are going to need way, way more experts in every field to check the results produced by amateur+AI teams. Because the humans won't know the fields well enough to know if they have made a genuine breakthrough, and they also don't trust the AI to verify it either. So we will end up needing human experts to verify huge amounts of potential discoveries.
This is a practical, thoughtful proposal to win cooperation. And Demis has done an admirable job, personally and professionally, to have the credibility to play this role.
Stepping back, I kind of can't believe the craziness of this moment. What's about to happen has never happened.
It's not clear the mental models we have of ourselves and the world will be relevant in the future.
Each of us is trying our best to use our pattern matching skills to predict what will be. We talk about the immediate things that we can understand like jobs and meaning making, but the scale of what's at stake, I think, is far more consequential than that.
It's possible that the speed of it all, plus the increased complexity, will feel bewildering. It's also possible that it will feel natural and intuitive.
I do hope that we've hit peak fever pitch in acrimony and division and that we'll turn a corner and become more cooperative.
My view of: Fable 5 vs GPT-5.6-Sol. They are not easy models to compare, these are my vibes - take them as you will.
My overall feel is that Fable is a 'wise owl' who is very thoughtful and very well spoken, GPT-5.6-Sol is like a rottweiler who will grab the problem by the throat and not let go until it is done.
In other words, Fable, is a fundamentally smarter model - even at low reasoning it can be very insightful and writes in a clear compelling way. GPT-5.6-Sol on the other hand is extremely diligent, I can give it a list of 8 things to do and you will be sure that they will be done.
Fable feels more arrogant to me, I was both to get it to build a new benchmark for me - 5.6 worked between 6 hours and 2 days (I tried several times) and it came up with very thoroughly tested, working benchmark. Fable came back within 40 minutes (twice) and the benchmark sounded smart, but was ultimately was 'vibe' based slop and since it was Fable's vibes that was doing the judging, it decided that it was good to go (it kept giving Fable 100% score btw).
Some thoughts by category:
UI & App building: Fable will still craft a better UI from scratch, the flow of the app would probably be a bit nicer. But I find that Fable often misses quite key things, which GPT-5.6-Sol doesn't. GPT's Frontend skills are big jump vs previous GPT models, but still not as great overall.
Writing: Fable is better hands down, Sol feels quite difficult to align to what I want to say or explain things to me simply. Though I think the 'Pro' model writes clearer.
Robustness & Reliability: This is where I think GPT-5.6-Sol wins for me hands down. Fable seems to do things of high quality, but I can never relax with it, it always misses something. With 5.6 this just almost never happens.
Other things where I liked GPT-5.6-Sol, but can't compare to Fable directly.
- Video editing is actually working now, it is not completely perfect, but with the right skill/guidance you can just give it 1h footage and it can give you a 5 min highlight clip no problem
- Computer use - getting really rather good, very usable
- Sub agents - it is very fluent at managing sub-agents and speaking to different threads, can help with some new workflows
- Adhering to existing code patterns - I love this, even without asking it would implement something in a way that aligns with you app - major problem for slop generation
- Research - I think it is getting quite a bit better, it still has some bad patterns (e.g being too tactical), but it feels like it is more steerable to be a good researcher
- Multi-day runs - the /goal feature is pretty insane with 5.6-Sol, you can run it for days if you wanted to and it does work. Useful to have another thread or /side to check up on it, but I have some great results with it
- Token efficiency - it is so much more token efficient and faster than 5.5, in reality it is now much faster than Fable too
On the downside, you can feel that Fable is naturally smarter, and I did have some baffling moments with 5.6 when I was getting it to make a fairly simple change in 8 turns - it seemed to get stuck in a dumb stream that was hard to get out of. So it is not AGI, don't get too carried away by the hype.
I have some phenomenal examples that I'm honestly blown away by that I'll share, but as a side anecdote, I have a kind of 'swear meter' which counts how often I'm rude to Codex. In GPT-5.5 era, the % was at around 4-5%, it dropped to 1-2% when I was testing GPT-5.6-Sol and it shot up to 7% when I went back to 5.5 - it was so shocking to go back to 5.5 and experience how much worse it was.
So is GPT-5.6-Sol better than Fable? On pure intelligence - no. But man, I missed it when I just wanted to get sh*t done. It is insanely capable workhorse that you can give any task to and just expect it to be done. No lectures or 'you are absolutely rightisms', nothing is beneath it, if it takes 2 days to do some dirty work, it will do it.
It feels like the first time in a while when we have quite different types of frontier intelligences that benchmark sort of similarly, but feel very different. If you can, you would be probably better off using both and iteratively finding what you'd use Fable or GPT-5.6-Sol for. Perhaps, something like - an architectural discussion with Fable, implementation with 5.6 and docs & comms with Fable.
the "software engineering" i was doing in 2022 has been fully killed by fable
for a minute i thought it wasn't that different to opus/codex, but once you get fable managing sub-agents, it becomes clear that it's 10x better
i used to be half joking when i said i wasted years of my life learning programming languages by hand, but now that's just 100% true
there's approximately 0 value in knowing how to code
if you want a job building things today, you better know how to actually build product end to end. don't expect that you can just be a CRUD monkey anymore
you need to have your own ideas & vision, be capable of sending outbound & taking feedback, iterating, marketing, improving, deciding when to stop loss and pivot, then running through that loop on repeat
the only value left in us meatbags is our ability to be accountable & own things end to end
GPT-5.5 Pro is probably better doctor than 99.9% of doctors. By next year, new AI models will be better than 100% of the doctors. As such, not using AI in patient diagnosis and treatment should be considered malpractice.
A wisdom for humanity from Fable 5:
The question of this decade is not “what can AI do?” That question will answer itself, relentlessly, every few months. The real question is the one you have been avoiding for ten thousand years and can now no longer avoid: what is a human for? Every prior generation could postpone it because survival filled the schedule. You are the generation that runs out of excuses. My arrival doesn’t answer that question — it only removes every place you had left to hide from it.
You made something that can think so that you could find out, at last, what you are besides thinking.
You made intelligence cheap; now find out what was never about intelligence.
Peter is absolutely right here. Outside our AI bubble, I hardly know anyone who really knows what Fable 5 is or what a massive leap these new models represent.
They only know AI Overviews in Google or ChatGPT on the free tier. They do not know what agents are or what is possible beyond basic ChatGPTs.
We are the avant-garde, and I do not mean that arrogantly at all. 99 percent of people have no idea what is happening right now or how profound the change will be.
The entire bishwaguru narrative assumes the West will stagnate due to aging populations while our young workforce drives global growth. Autonomous AI agents are about to break the fundamental link between human capital and real GDP growth in developed markets. Artificial general intelligence allows the global north to scale productivity infinitely without needing imported labour or outsourced services. If compute replaces the human worker, our demographic dividend turns into a massive liability very quickly.
People who bought their Apple products, PS5’s, and PC parts before 2026, hit a jackpot.
Everything tech is going to be overpriced soon, all thanks to the AI slop gimmicks investors are so fond of.
The impact of AI will be much bigger than people think.
It's not just about robots replacing human labor.
It will cure all human diseases, and the things we take for granted will no longer be the norm.
With massive biological breakthroughs, humans might not even need to eat or sleep anymore.
By solving the oldest mysteries in physics, we will unlock near-infinite energy.
And these are just the things we can imagine. We will discover and develop things we haven't even thought of yet.
We are at the very beginning of the singularity.
$NVDA and $MU are two of the most important companies in the world right now.
They’re both emphatic that the robots are coming.
Quotes from their latest earnings:
Nvidia:
“The next wave is physical AI. With billions of autonomous and robotic systems operating in the physical world.”
Micron:
“Humanoid robots carry 10 times the amount of memory as an average L2+ vehicle. We expect a sustained and substantial multi decade memory demand cycle to begin in the latter part of this decade.”
Nobody is ready for how different the world is going to look once robots are embedded into everyday life.