It was the best of times,
it was the worst of times,
it was the age of wisdom,
it was the age of foolishness,
it was the epoch of belief,
it was the epoch of incredulity,
it was the season of Light,
it was the season of Darkness.
From A Tale of Two Cities, by Charles Dickens
Actually, you could also run a subagent in the main thread and call tools like kimi-cli to use the k3 model in max mode. I usually do this for adversarial review of solutions.
Let Sol manage an efficient fleet of Luna agents for you.
These models know each other well and collaborate to achieve the result in an incredibly fast and efficient way.
Announcing Discovery Loop!
I am very excited to announce that, along with my longtime friends and collaborators @Sanjay_Ghemawat, @OriolVinyalsML and @quocleix, we are founding Discovery Loop (@DiscoLoopAI), a Public Benefit Corporation whose mission is to automate machine learning, science, and engineering to accelerate discoveries and progress. The four of us have worked together for 14 to 30 years, and have helped build some of the world’s most used products, infrastructure and AI models, and we’re excited to turn our attention to this ambitious endeavor.
♾
Learn more at: https://t.co/Rv3LMdLluK
Oops... I did it again.
Enjoy reset usage limits for all paid users for Codex and ChatGPT Work. Super grateful for an incredible team who is iterating at lightspeed and keeping the infra up as we scale faster than ever.
Enjoy the weekend!
To celebrate the launch of GPT-5.6 Sol, we will reset the rate limits again (twice) across ChatGPT Work and Codex over the next 24 hours.
We want you to have the time to truly try ambitious tasks and get the hang of it. Happy exploring!
Hi. Over the last 24 hours we had three separate small incidents that affected Codex reliability. Those are three too many and we are taking active steps for them to not reproduce.
I have reset usage limits for Codex across all paid plans. May the tokens flow again.
Personal update: I've joined Anthropic. I think the next few years at the frontier of LLMs will be especially formative. I am very excited to join the team here and get back to R&D. I remain deeply passionate about education and plan to resume my work on it in time.
Do not use Codex as autocomplete with a shell. For complex work, start with /goal <objective> and make the stopping condition explicit: which user path, system behavior, or deliverable must become true, and which tests, builds, logs, screenshots, or API probes will prove it. Then use Plan Mode to read the repo and narrow scope, grill-me to expose hidden dependencies and failure paths, SDD to turn the plan into a contract, and TDD to create red/green feedback. /goal is not a do-my-ticket button. It is a constraint workflow. Without evidence, `done` is just a confident summary.
You've been asking for this one...
Now in preview: Codex in the ChatGPT mobile app.
Start new work, review outputs, steer execution, and approve next steps, all from the ChatGPT mobile app. Codex will keep running on your laptop, Mac mini, or devbox.
New on the Engineering Blog:
Building Managed Agents—our hosted service for long-running agents—meant solving an old problem in computing: how to design a system for “programs as yet unthought of.”
Read more: https://t.co/YYaEub2QGV