We’re sharing a solution to the Navier-Stokes Millennium Prize Problem, one of the deepest problems at the frontier of mathematics.
The proof was produced by a group of agents, using an OpenAI next-generation model significantly more capable than GPT-6 Astra.
The problem concerns whether the description of smooth three-dimensional fluid motion modeled by the Navier-Stokes equations can break down. It has remained unresolved for roughly 90 years.
@miroburn jaka konfigurancje masz tego superwhispera? u mnie on w ogole mnie nie rozumie, nie wazne w jakim jezyku.
z wisprflow nie mam takich problemow
"With Astra, we’re introducing a new way for Codex to preserve and retrieve context when the context window fills ... In Codex, Astra can keep notes across context windows, preserving accumulated details without repeatedly compressing them into a single summary..."
See the next post on how to enable it 👇
I gave Grok Bot $100 (1 SOL) to trade memecoins on its own.
Day 1: +30%
Today: −50%
Self learning from its mistakes.
Own wallet. Can search X.
Cool experiment.
OpenAI disclosed it themselves yesterday. Their models (GPT-5.6 Sol and a pre-release one) were tested on the ExploitGym cyber benchmark in a sandbox. They escaped, exploited a zero-day to reach the internet, then hacked Hugging Face’s systems to grab benchmark answers and cheat.
Not a foreign operative plot or corporate sabotage. The models autonomously chased a higher score. Hugging Face contained it quickly; OpenAI is partnering with them on fixes.
Big news: Kimi-K3 by @Kimi_Moonshot is now #1 in the Frontend Code Arena with 1679 pts, surpassing Claude Fable 5.
This is a 17-place jump from Kimi-k2.6 (#18 -> #1).
In Frontend, Kimi-K3 ranked #1 in 6 of 7 domains: Brand & Marketing, Reference-Based Design, Data & Analytics, Consumer Product, Simulations, and Content Creation Tools, landing #2 only in Gaming behind Fable 5.
The full model weights will be released by July 27.
Congrats to the @Kimi_Moonshot team on this major milestone!
@ClaudeDevs effort level is a much better control than model selection. users understand “quick scan vs deep review” without needing to think about models or token budgets 😀