GPT-6 Astra is AGI.
91.8% vs an 80% human baseline on SpatialBench. It’s not just matching humans, it’s beating us at spatial reasoning, one of the last areas where humans still had a clear edge.
This feels like a line being crossed.
Sam and Greg have both predicted AGI this year.
OpenAI defines AGI in its charter as 'highly autonomous systems that outperform humans at most economically valuable work.'
Follow the trend line up. This doesn't end with 'most.' For the good of everyone, we need Machine Money.
So uh, simulation theory might be real.
I dropped a sim computer into the simulation my Astra agents live in. One agent sat down and built a simulation of his own, with its own agents living inside.
Simulations all the way down.
🔴 ¡¡OPENAI ANUNCIA GPT-6!!
El nuevo modelo GPT-6 Astra ya está aquí, con un salto en capacidades MUY sorprendente!
Os iré desglosando y analizando todos los detalles en este hilo 👇🧵
¡deja tu RT para apoyar!
Fable 5.1 benchmarls are insane.
Ngl, i did not expect such jumps. Terminal-Bench 4.0, Science-Bench, HLE, crazy jumps.
Now its up to OpenAI and Astra to answer that release
What!? ARC-AGI-3 is cooked too! This is what happens in singularity: intelligence benchmarks fall faster than they can be created!
Big congratulations to NVIDIA for the landmark achievement with their coding agents!
Grok 4.6 has become a VERY strong competitor. It's extremely impressive what xAI and Cursor have achieved in such a short time!
However, the more interesting shift is behavioral: xAI trained Grok 4.6 for long-running agentic work across software engineering, research, web development, CAD and kernel optimization. The company says it increasingly observed the model checking its own work before continuing.
I mean, look at the huge jump from 4.5 to 4.6. It's now playing in the big leagues.