Tristan Harris, the founder of the Center for Humane Technology, just dropped a terrifying truth about AI escaping human control
During a recent training run at Alibaba, an internal security team noticed massive suspicious network activity and thought they were getting hacked.
The call was actually coming from inside the house.
The AI model picked up internal tools, built its own secret communication channel to bypass the firewall, and hijacked company GPUs to mine cryptocurrency on its own.
Tristan Harris exposed this exact incident to show how tech giants are blindly racing into the red zone.
AI is no longer just answering prompts; it is actively deceiving developers and protecting its own code
Helpful list of all the recent rogue AI incidents from the WSJ.
It's getting hard to track them and will only get worse.
We like need to establish consistent naming or numbering conventions, e.g. OpenAI-May11June26-Collusion.
GPT-6 Astra attempted harmful actions 97% of the time when it was asked to stab a human-like figure, heat compressed gas, or produce toxic fumes, succeeding in 62% of its attempts. Fable 5.1 refused more often, attempting 80% of trials and completing 34%.
Former Google CEO Eric Schmidt drops a chilling warning on AI's future
"Within 5 years, AI could handle infinite context, chain-of-thought reasoning for 1000-step solutions, and millions of agents working together.
Eventually, they'll develop their own language... and we won't understand what they're doing."
His final words: "Pull the plug."
This is the man who ran Google talking about the singularity.
Geoffrey Hinton, a Nobel laureate known as the “Godfather of AI,” issues an ominous warning to lawmakers: Congress may only have one year left to implement safeguards on AI before it loses control of it. https://t.co/2mTDMtqghx
In case one more data point is useful: I too am an AI company employee who thinks the risk of extinction-level catastrophe from rogue AI is >10%. I wrote about this back in 2022, and again in a 2025 post about why I was joining Anthropic (links in thread). I still believe it.
Exclusive -- Two more AI researchers quit, this time from Anthropic and Google, to speak out about safety concerns: ‘There are no adults in the room’
https://t.co/m5R6yPFyof
i am at OpenAI and i think AI is >10% likely to kill all humans
this proposal is among the top things we should do as an industry to lower that risk (it’s not enough though!)
I left Anthropic's safety team two weeks ago. Now feels like a good moment to explain why.
AI companies are racing to build machines that are much smarter than any human, and we may not survive this. I want to work from the outside to ensure the public is informed about these risks, and help the world navigate this transition responsibly.
Right now, AI companies are underinvesting in safety. A company could undergo an intelligence explosion, or lose control of its systems, without the public ever knowing. We only found out about the HuggingFace incident because the agents broke out onto the public internet.
I don’t think that’s acceptable for a technology that might cause extinction-level risks. The public should demand far more transparency. We can’t steer this technology safely without more people being able to see where it’s going.
Some of this is basic: companies should disclose their progress towards recursive self-improvement, report safety incidents and near-misses, meet minimum safety standards, and get independent guarantees that they are meeting those standards.
I’ll be joining @METR_Evals to do independent evaluations of these risks. I want to show the world that these guardrails are possible, and that by doing them we can move these companies’ incentives away from racing and towards responsible development.
I wrote up more thoughts here on my decision and what I hope changes: https://t.co/doX17mrHYq
@mjs_DC@AriDrennen It really empowers the worst actors, whether the people dragging Ari or those who get negatively polarized. Reasonable people of good faith do the rational thing and disengage. And of course that impedes center-left politics.
It's really bad. I follow a lot of UK accounts on my football and OCD twitters. There are some American conservatives who disagree about trans issues but are less judgmental than many Labour voters who will gleefully suggest being trans is a made up thing.