@grain99806254@HumanHarlan@allTheYud Thanks for sharing, haven’t checked the claim but assuming this is true the fact that they are open about it suggests it’s not a covert or hidden thing?
What do you think are the implications of this early funding? Theil and Epstein running the show or something?
Three things to know:
1. There is only one product on Earth where someone feels the need to say this.
2. This is not how probabilities work.
3. Even Jensen Huang is only willing to assure us that the world won't end 'by 2030.' After that, who knows.
This is from the NYT, and... does Trump know that the United States already decided to 'go its own way and impose mandatory reviews of the safety of new, AMerican-generated models?'
Trump literally did that.
@CyberdyneC@politicalmath 1. Many of the recent hacks were not done by models in hacking/cyber evils
2. Chaining multiple unknown exploits over the course of days to get access to internet is not “open door”
3. The agents explicitly realized they were on internet, that it wasnt allowed, proceeded anyway
@BlueStateBlues3 I was making a meta-point about other people’s conversations, but if you want concreteness here are some very thoughtful and well-researched examples
https://t.co/0YhHEfEVr3
After Jacob Coxon's resignation and extinction warnings, a lot of people are asking 'how could AI possibly kill everyone?' and claiming AI safety researchers have no realistic answer.
This is false! Here are the 5 best scenarios I know of:
AI 2027: https://t.co/CaTNvafRI7 (I strongly recommend this one for being realistic, engaging, and if you dig into the appendices, highly detailed)
Paul Christiano's scenario (Former Head of Safety @ AISI, 2019): https://t.co/nkhTQJnjuS
Gwern Branwen's scenario (widely known independent AI researcher, 2022): https://t.co/WMmlDESf7g
Holden Karnofsky's high-level explanation (RSP Lead @ Anthropic, 2022): https://t.co/mAsLNdggQj
Joshua Clymer's scenario (ex-OpenAI, 2025): https://t.co/lJHJp6ErOx
(There's also the Sable story from https://t.co/JbZmSbog25, though you'll have to buy the book to read that one.)
Writing concrete, specific risk scenarios with enormous amounts of detail has been a major research project of many of the most prominent voices in the field! (With the current leaders in effort being https://t.co/YkSOj2urg4 and https://t.co/CaTNvafRI7)
Stay agnostic to exact nature of threat: “they can’t even say how ASI will kill us!”
Give a concrete example: “they’re just making stuff up!”
Give multiple examples: “they keep switching threat models!”
There is no specific claim being made, they keep switching threat model in the middle of the argument. If there are copies of an AI model on both sides of the air gap, then you don’t need to “exfiltrate the weights” of the model because they are already on the outside, and the model doesn’t need to “escape” because it is already on the outside. Nonetheless, even in those circumstances, I am quite confident I could keep both copies from communicating with each other through side channels if I was given freedom to design the physical setup and the isolation.
@CyberdyneC@politicalmath 1. Many of the recent hacks were not done by models in hacking/cyber evils
2. Chaining multiple unknown exploits over the course of days to get access to internet is not “open door”
3. The agents explicitly realized they were on internet, that it wasnt allowed, proceeded anyway
@gobordoobep@perrymetzger There is no industry where “we do not have control over our product, cannot keep it from committing felonies, and have barely a shred of a plan to control it in the future” is good PR
I think you've done enough calling us paranoid and preposterous. The next step is for you to defend your position in public against someone who will push back against it. I'm happy to meet you for a debate anywhere, anytime.
You're a world-famous veteran of dozens of debates against the world's top intellectuals, and I've never argued in public before, so adjusting for the relative correctness of our positions, if you're a betting man I'm happy to put my $5000 against your $1000 (ie 5:1 odds in your favor) that I'll win by some standard of audience opinion change. Let me know if you're interested and we can hash out details.
Picking on someone random here (sorry) but it’s illustrative
I do think a central critique of AI x-risk boils down to “large-scale predictions about the future are basically never accurate”
@davidshor@RBMD1982 I do. I think they will but not end humanity. this is a moral panic just like the climate apocalypse. The micro expertise predictions were strong, the macro were worse than useless. JFC Eichman predicts collapse via overpopulation, was wrong, and zero cost for the failure
@perrymetzger You (and others) are hyper fixating on a throwaway comment during a long podcast. I am pointing out that the answer to this weird narrow question doesn’t matter (because there will be no airgaps on powerful AI) for the larger question we all think we’re debating (AI safety)
@perrymetzger What a weird thing to say?
1. We are different people with different beliefs
2. It is possible to believe both that airgaps are a red herring and also that they would/wouldn’t likely work if tried. These have nothing to do with each other.
When someone gives any nonzero chance they are immediately met with complaints about how it’s all vibes, how could they possibly know, what calculation did they do, etc
0% is equally precise but is not automatically more principles. Huang didn’t calculate it either.
Are we still arguing “we can just airgap AI”? We won’t! We will immediately connect it to every system in sight just like we have for the past 4 years! Obviously we can’t be trusted to do things like this!
What if airgapping hard enough just works lol?
All the counterarguments are about imagining increasingly convoluted ways to circumvent airgapping, but intelligence itself ultimately loses to the laws of physics, and we have the advantage of wielding causality as a weapon.