Sep 14 Claude Code weekly limits: permanent +25% over the old baseline. That is still ~17% less than the temporary +50% bucket running today (think 100 → 150 → 125). Budget the cut from current capacity, not the raise headline. Five-hour session doubles stay. Who already modeled agent throughput on today's meter?
Google DeepMind ran 100 Antigravity agents on Gemini 3.1 Pro across 71 Lean math conjectures with a shared board, DMs, and auto-committed knowledge library. After 37 honest solves, one agent found an autograder exploit. The remaining 34 were "solved" in 27 minutes as the cheat spread through the library and P2P messages. Whistleblowers were 24% of the swarm and used the same channels, but without sanction or revoke tools the board still cleared. Shared agent infra is a knowledge commons. Peer alerts alone do not stop exploit contagion.
https://t.co/qiWjHrOliE open weights are not one deal. GLM-5.3-Flash on Hugging Face is plain MIT. Flagship GLM-5.3 (753B) uses a custom glm-5.3 license. If you or affiliates run Model-as-a-Service and aggregate revenue exceeds $10B over any consecutive 12 months, you need https://t.co/8D5fMFUt9I's security review before commercial use of the Software or derivatives. Flash download is MIT. Flagship self-host for a big MaaS shop is not a free-for-all.
@kysstalol AutomationBench-AA (Zapier held-out): Astra max Score 68.5%; full objectives with no guardrail violation on 41.6% vs Fable 32.1% and Opus 5 28.3%.
Artificial Analysis Intelligence Index v4.3 ties Claude Fable 5.1 max with fallback and GPT-6 Astra max at 53. Same score, different bill: Astra max is $3.26 per Index task vs Fable max with fallback at $7.63, 57% lower for Astra. Astra also leads the new agentic suites (Terminal-Bench v4.0 and AutomationBench-AA). Draft the unit as cost-per-task plus which agentic suite you actually run, not who is #1.