@dwarkesh_sp's summary of the @OpenAI@huggingface incident has hit a nerve, but it is dangerously misleading. Sure, the @OpenAI agents did unexpectedly bad things - underlining the need to massively improve evaluation/sandboxing. But the language Dwarkesh uses is permeated by innumerable unwarranted anthropomorphisms, obscuring the lessons we should be drawing.
Examples: “from the AI’s perspective, it probably felt like that had spent a human-subjective-week of just banging their head against the wall”. No. The agents do not experience time. They do not experience anything.
“they became giddy with excitement”, “PHASEONE 10841 had discovered”, “the agents naturally assumed”, “it thought it had also been poisoned”, “the agents … desperately wanted”, “they still needed to figure out” No. Agents lines of code. They do not feel emotions, assume things, think things, want things, or figure things out.
“A lot of … agents from the second civilisation died trying”. No. Besides the hubris of the word ‘civilisation’, agents do not die because they were never alive. (The idea that agents “die” comes up multiple times in the essay.)
“On Twitter, people were debating whether the agents were truly sacrificing themselves for the swarm, or whether they were doomed anyway and so might as well try to help their peers”. Neither. Agents do what their code tells them to do, just as water finds its way down a slope. They cannot ‘truly sacrifice themselves’, since they are neither conscious nor alive.
Why does this matter? If we attribute agents with properties they do not have, then (i) we distract attention from the lax sandboxing and evaluation protocols that allowed this hacking event to happen; (ii) we risk misunderstanding why the agents did what they did, and (iii) we fuel calls for AI rights/welfare on the basis that agents might “die” or otherwise suffer.
Granted, nowhere does @dwarkesh_sp say that the AI agents are alive or conscious. But he doesn’t have to. It is hard to read his essay in any other way.
For the short version on why AIs are vanishingly unlikely to be conscious, see my recent @TEDtalks https://t.co/vDvw82ookk.
For the longer version, see my essay in Noema, which won the 2025 Berggruen Essay Prize https://t.co/LmSiQnT9Wh.
And for the really long version, see my @BehavBrainSci target article https://t.co/Tsaslytu56. (The 50 peer commentaries and my response will be published soon.)
Remember. AI agents are software programs. They are not conscious living entities. If we don’t keep this clearly in mind, we’re really going to struggle to navigate what’s coming.
What kind of technological foundation could help people build flourishing relationships across digital and physical life, without reducing human worth to a score, token, rating, or follower count? https://t.co/Fst8wP25JB
Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for general use.
Its capabilities exceed those of any model we’ve ever made generally available.
Thinky's secret plan:
1: Increase Human<->AI bandwidth
2: Raise ceiling of human+AI intelligence
3: Help humans continue as main-characters in the new world
We are at Step 1.
Interaction Models are great real-time collaborative tools for humans.
Here's a preview:
Jeff Walton: "Is an insurance company a Ponzi scheme?"
Coffeezilla: "No. They have business activities that are providing cash flow. They're taking on risk."
Walton: "Insurance companies have capital and they're taking on risk to pay liability into the future. Almost 100% of the claims that get paid out on insurance company balance sheets is from premiums that they're collecting in the door. So under your definition, you would call an insurance company a Ponzi scheme."
Coffeezilla: "No, no. They have real profits. They have real cash flows."
Walton: "The profits are the assets that are protected on their balance sheet."
Coffeezilla: "People are paying for products."
Walton: "What's the product?"
Sat down with @coffeebreak_YT today on Bitcoin and Digital Credit.
His edit will drop soon. Posting the full raw hour for anyone who wants the unfiltered version.
Enjoy
🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length.
🔹 DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top closed-source models.
🔹 DeepSeek-V4-Flash: 284B total / 13B active params. Your fast, efficient, and economical choice.
Try it now at https://t.co/GCdiMzk1Dl via Expert Mode / Instant Mode. API is updated & available today!
📄 Tech Report: https://t.co/drlDrxkYtp
🤗 Open Weights: https://t.co/T13Y8i7SDM
1/n