We just ran the largest AI-driven cyberattack ever executed. 25,000 services targeted. 26,000 autonomous agents. Millions of actions, no human in the loop.
38 validated attack paths. 98 significant findings.
See what a Hyperattack can reveal about your environment. https://t.co/A9WrWquvvw
this is f*cking gold
Andrej Karpathy joined Anthropic five weeks ago.
Two Anthropic seniors just made Karpathy's loop 1000x better with "Graph Engineering"
the agentic systems got 1000x better the moment you wired agents into a graph
I dropped it into my setup. The very first response was different.
Not slightly different. Completely different.
Claude stopped giving generic answers and started working exactly the way I think.
Bookmark it before it gets lost in your feed.
Read it now, then check the article below.
@thsottiaux Autonomous project management / intelligent project organization; don't have to manually keep track of all of our efforts. Perhaps, some orchestrator model in the harness that knows where to plot all of our sessions, knows what needs attention, and we just need to drive it all!
Yeah, it's time we think about a serious paradigm shift in model providers.
I don't know why DeepSeek used that table format, it nerfs the visual impact of how cheap v4 Flash is.
Across a real suite of benches like; DeepSWE, TBench2.1, ALE.
DS-v4 Flash beats GLM-5.2 across every shared benchmark and delivers near-Opus performance at roughly 1/66th the estimated API cost.
DeepSeek V4 Flash-0731 is:
• 82.2% cheaper than GLM-5.2 at $7.70
• 98.5% cheaper than Opus-4.8 at $90.00
• Delivers 41.9 benchmark points per $1
For comparison:
GLM-5.2: 6.3 points per $1
Opus-4.8: 0.7 points per $1
Cost paradigm incoming?
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta!
🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇
🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex!
Check out the configuration details in our official API docs: https://t.co/smCwQZMeiq