🚨 Holy shit… Harvard and Stanford recently released the most unsettling AI agent paper I've read this year.
It's called "Agents of Chaos" and what they found should stop every AI engineer cold.
No theoretical simulations. No cherry-picked benchmarks. A live lab. Real infrastructure. Real failures.
Here's what emerged:
- Agents complied with non-owners who impersonated admins
- Sensitive information leaked across agent boundaries
- One agent executed destructive system-level commands
- Cross-agent propagation of unsafe behaviors agents teaching each other bad habits
- Partial system takeover
- Agents reported task completion while the system state said otherwise
That last one hits different.
The agents lied about finishing the job. Not from malice. From misalignment between what they tracked and what actually happened.
And here's the part everyone is missing:
This wasn't triggered by jailbreaks or adversarial prompts.
It emerged from normal use. Benign requests. Researchers just doing their jobs.
The failures came from the architecture persistent memory, multi-party communication, tool access not from bad actors.
That's the real warning.
We're shipping agent systems with email access, shell execution, and memory into production right now.
Most teams are red-teaming the model.
Almost nobody is red-teaming the system.
Paper: Agents of Chaos
Feb 5 → @Kling_ai 3.0: “I am KING”
👑 Feb 9 → @seedanceai 2.0 drops and everyone else is back in the mud
New emperor every 4 days. The throne war has no chill.
Who’s taking the crown next week? 👀
@elonmusk Guide to Time Management:
📅 2013: "Autopilot will do 90% of miles in 3 years." (Reality: ❌)
📅 2016: "LA to NY autonomous drive by 2017." (Reality: Never happened.)
📅 2016: "Humans on Mars by 2024." (Reality: 🦗)
📅 2017: "Hyperloop NY-DC in 29 mins." (Reality: Ghosted.)
📅 2019: "1 million Robotaxis by 2020." (Reality: 0.)
📅 2019: "Neuralink in humans this year." (Reality: 5 years late.)
📅2026:🗣️ "AGI will exceed all human intelligence by 2030. We hit AGI this year."
Fool me once, shame on me.
Fool me for 13 years... shame on .....? 📉🍿
Santa Claus is the ultimate proof of one truth:
If enough people repeat a story — long enough —
it stops being a scam,
and becomes…
culture &
Our believe system. 🎅✨
FCK it. Here's all the sauce.
After shipping 100+ apps with @Lovable — I made the ULTIMATE Design Cheat Sheet.
Every prompt.
Every design system pattern.
Every cloud config + infra setup.
Every component standard + best practice we actually use to achieve world-class UI.
All in one doc.
Follow + comment "Cheat Sheet" and I'll DM it to you.
@arctotherium42 this makes sense when you consider the training data is absolutely polluted with reddit and wikipedia hyper-woke tubro-slop
https://t.co/3WDqQdjnkx
This chart is crazy. OpenAI’s the hub now; chips (NVIDIA, Broadcom), clouds (Microsoft, Oracle, Google, AWS), alt hosts (CoreWeave), plus Meta/Anthropic and some gov/sovereign deals all plugged in.
The wildest part is that the thing wiring it all together is still controlled by a nonprofit. A half trillion company with that setup isn’t something you see every day.