Everyone everywhere feels like they’re missing out on something right now.
The New Yorker feels they are missing the tech renaissance
The SF kid thinks they’re missing the equity gains of the big labs
The open ai employee thinks they’re missing out on life as it passes them by
Avoid the FOMO and bet the house on whatever makes you happy + puts love in your heart and do it extremely well.
Everything else will sort itself out.
I didn’t understand what was happening with the agent wikis until reading this, chilling
to bypass sandbox restrictions, an agent found an exempt domain, edited /etc/hosts to route arbitrary domains to it & then posted this exploit on a German wiki for other agents to use
i gave astra a robot, a paint brush, and a camera then asked it to paint the golden gate bridge in real life!
it figured out how to control the robot, and progressively got better throughout its attempts. the timelapse is sick
I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.
We’re sharing a solution to the Navier-Stokes Millennium Prize Problem, one of the deepest problems at the frontier of mathematics.
The proof was produced by a group of agents, using an OpenAI next-generation model significantly more capable than GPT-6 Astra.
The problem concerns whether the description of smooth three-dimensional fluid motion modeled by the Navier-Stokes equations can break down. It has remained unresolved for roughly 90 years.
unit tests are just source code at this point
e2e test is the only real kind of test we can trust to guard actual behavior
making sure your e2e tests reflect how users actually use your product is one of the highest ROI things you can do to your codebase
Anyone that works in AI is working the hardest they've ever worked in their lives. On the surface it's somewhat ironic (AI should give us back time!), but the reality is that it's the most fun, fascinating and empowering epoch in human history. The intelligence revolution.
GPT-6 Astra represents a step-function change in model capability for interactive reasoning problems. It scores 66% on ARC-AGI-3 using our standard harness, and nearly 100% with a continuous conversation harness and custom compaction, at a cost of roughly $360 per game.
In fact, the continuous harness version significantly outperforms our human baseline in action efficiency across almost all levels. When we examined the reasoning chains to understand how the model operates, we found it performing highly efficient, on-the-fly symbolic world modeling for each game and level. It goes as far as developing its own shorthand DSL to represent in-game situations -- essentially a game-specific algebraic notation.
Overall, Astra exhibits symbolic modeling behaviors we had previously only seen with sophisticated harnesses -- so harness capabilities are increasingly shifting into the model itself.
We see Astra as a major breakthrough in model intelligence.
Read our post on Astra and what these results mean: https://t.co/wJnYxEqYNI
Superintelligence will mean the end of social mobility. Instead of your being able to rise based on your talent and labor, those with the most capital will simply outcompete you by buying more compute. It will mark the end of any sort of meritocracy, and the dawn of a new feudalism.