An internal version of Astra, @OpenAI’s next major model family, solved 10 major open problems in mathematics, quantum complexity, and theoretical computer science.
We believe it will be a major step for scientific reasoning. https://t.co/iP6cyheZ7i
🚨 OpenAI preparing to release a new model family, tentatively named "Astra"
OAI touted its abilities to have multiple agents work together over a long period of time to solve particularly hard problems
Astra is a new class of OpenAI models alongside Sol, Terra and Luna
h/t @bughuntergeek
Claude and codex have gotten to the point where they invent their own project specific vocab. I use the Simplified Technical English prompt often when it gets out of hand.
That seems to reasonate, the community put together skills and guides based on that idea
Claude skill:
https://t.co/vFecMjo9MQ
STE skill kit:
https://t.co/Qi8rxXWXPD
Writing styles:
https://t.co/xXjkp6Jv8A
Runbook skill:
https://t.co/j5MHAaz3Kz
STE checker:
https://t.co/XnUmPLuvCG
Enjoy! Happy slop combat
The AI price war is ON.
> DeepSeek just dropped V4 Flash 0731 (big capability jump, $0.28/M output)
> OpenAI cut GPT-5.6 Luna by 80% to $1.20/M (OpenRouter is currently offering 50% off, bringing it to $0.60/M)
Same budget tier now, so I ran both through 3 canvas tests:
🧊 Rubik's cube: DeepSeek nails every rotation. Luna's turns are buggy, stickers fly off mid-move.
🎆 Fireworks: both fine, but Luna lags hard when the burst explodes.
🖊️ Pen writing "Hello": neither is legible. DeepSeek's cursive is looser, but Luna gave up on animating, it rendered a static image.
GPT-5.6 is great at coding — but IMO only on Sol. In the same pricing tier, DeepSeek clearly beats Luna.
DeepSeek v4 Flash 0731 weights were just published!
Plus:
- technical report
- MIT license
"DeepSeek-V4-Flash-0731 is the official release of DeepSeek-V4-Flash, superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as DeepSeek-V4-Flash-DSpark, i.e. it comes with a speculative decoding module attached.
DeepSeek-V4-Flash-0731 outperforms DeepSeek-V4-Pro (Preview) on benchmarks listed below despite its far smaller activated parameter count, and is broadly competitive with the strongest proprietary models available."
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta!
🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇
🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex!
Check out the configuration details in our official API docs: https://t.co/smCwQZMeiq