The new GPT-5.5 Cyber model achieves a higher CyberGym score than Mythos 5
Among open-source models on this benchmark, GLM-5.1 is the top performer, scoring 68.7%, comparable to Claude Opus 4.6
I’d be curious to see how GLM-5.2 performs on CyberGym
We’re expanding OpenAI Daybreak to help democratize patching vulnerable software at machine speed:
- Codex Security plugin: find, validate, and fix vulnerabilities right inside Codex
- The full version of GPT-5.5-Cyber model: a great model for trusted defenders
- Cyber Partner Program: powering products built on top of our best cyber capabilities for leading security companies to secure the world's software
- Patch the Planet: working with maintainers to secure critical open source projects
https://t.co/hyIi6gQmkm
We’ve been researching new ways for ChatGPT memory to carry context across conversations and keep it useful over time.
Today, that work is rolling out as a more capable memory system in ChatGPT. https://t.co/0MyFKCe2Mu
You can use codex within your own programs using the Python SDK. It's awesome. Built by @ah20im and friends
```
pip install openai-codex
```
https://t.co/GjQVEPwtkF
Windows users, this one’s for you.
Computer use now works on Windows, so Codex can take action on your Windows computer.
And with Windows support for Codex in the ChatGPT mobile app, you can start, review, and steer tasks on the go while work continues on your Windows machine.
An early experience, but we’re working on more ways to keep your work moving, wherever you are.
New data from @NASAWebb shows that supermassive black holes can grow to their current size without a much larger host galaxy to feed them.
This helps to explain why some black holes in the early universe got so big so quickly. https://t.co/9l3SjnKHqZ
Builders Unscripted with @0xmts
Matias talked to @romainhuet about bringing Codex to work and into side-project workflows.
00:58 Codex at Alchemy
01:51 Code review catches bugs
08:04 Side projects with Codex
18:51 Codex App Server projects
24:01 Computer use, GPT-5.5, SnapCat
AI can give researchers the freedom to pursue “crazier” ideas.
For Terence Tao, AI creates more room to experiment, test unexpected paths, and discover what might otherwise stay out of reach.
Grok is very good at asking clarifying questions. Codex is not. Grok's queueing issues don't impact the specifier role of the swarm-forge very much. It's a bit less rigorous than codex in following rules, but I think I'll use it from now on in the specifier role anyway.