An AI agent recovered a quantum laser lock 99.3% of the time — no human.
Anthropic opened MHS — a preview so agents can run lab hardware.
> Setup: weeks/months → hours/minutes
> CMU dose-response ~3× faster
> QuEra lock recovery 99.3%
Would you let Claude run the bench?
@GoogleDeepMind@FlowbyGoogle 40 seconds is the new cap. 10-second steps, and it now reads the last 10s of the clip you already made, not just the last second.
Claude closed 85% of a deception gap. Six researchers closed 20% — same rules.
Anthropic had Claude search, train, and test 10 alignment failures.
> Held on models 4.7× larger
> Sonnet 5 → early Opus 4.8 in 60h · ~2k · ~15,000×
Would you let Claude align the next Claude?
The OpenRouter stealth #1 is now a MIT download — and a 57.
https://t.co/zBNCplx9D4 shipped GLM-5.3-Flash (ox-alpha): 320B, 18B active.
> AA Index 57 · $0.045/task discounted
> DeepSWE 63.4 vs GLM-5.2 46.2 (vendor)
> AutomationBench 48.8 vs 26.2 (vendor)
Pulling the weights?
The free mystery model that doubled DeepSeek's usage is a GLM.
https://t.co/vwpdYaIxxe confirmed Ox Alpha to Bloomberg. Weights drop tonight.
> #1 OpenRouter usage
> 2× DeepSeek · still $0
> New GLM · weights tonight
Pulling the checkpoint?