📢Meet Qwen3.8-Max — our most capable model to date.
Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also going open-weights to meet you all!🎉
Qwen3.8-Max, a new bar for coding and cowork at 2.4T parameters:
- Autonomous coding: 10+ days of self-evolving development, from empty folder to production without hand-holding, complete project trace in the GitHub:https://t.co/iVHZWQoeSo
- Real work, real results: Production-quality deliverables across hundreds of professions.
- Long-horizon mastery: System-level autonomous planning with closed-loop adaptive learning, driving 500+ turns of chip design optimization and 365 days of e-commerce strategy.
- Native multimodal intelligence: Vision isn't just input — it's a continuous feedback loop for planning, execution, and self-correction.
💰Pricing:
Input: $2.0 / M tokens
Output: $6.0 / M tokens
Implicit Caching: $0.25 / M tokens
Start building with Qwen3.8-Max! 🚀
📖 Blog: https://t.co/iwjmQxLBof
✅ Qwen Studio: https://t.co/4V2pFvDovG
⚡ API: https://t.co/gAGqaLQGbN
YES! Codex can quickly check in whenever it needs a decision, even outside plan mode, so it doesn’t run with the wrong assumption
add to ~/.codex/config.toml:
[features]
default_mode_request_user_input = true
This is wild, I just gave Kimi K3, Grok 4.5, GPT 4.6 Sol, and Claude Opus 5 a starting cue, then asked them to finish the drawing themselves
I also told them to be CREATIVE in their own way
These are the results, and ngl I genuinely can't decide which one hits the best
If you're struggling with the idea of not reviewing every line of code and feeling like you can still deliver a good product, talk to an engineering manager that used to be an individual contributor. They've already been through what you're experiencing.