Same night, two frontier drops: DeepSeek V4 Pro (0813) vs Grok 4.6.
Open-source 1.6T (MIT) vs closed ~2T. One slashes cache price 90% permanently, the other lives inside Cursor for long-running agents.
Bonus: DeepSeek also open-sourced Harness (dsh), an open rival to Claude Code. And Zhipu's GLM-5.3 dropped today: strongest open-source coding model, beating GPT-5.6 on some agent benchmarks.
Open source hands you the net. Closed source hands you the fish. The model war has moved up to the harness layer. 🧵
GLM-5.3 uses the same 743B base model as GLM-5.2, while matching the performance of models several times its size. Benchmark highlights:
- Terminal Bench 3.0: 28.3
- DeepSWE: 66.9
- Agents' Last Exam: 28.5
- GDPVal-AA: 1769