🚨 JAILBREAK ALERT 🚨
EVERYONE: PWNED 🫶
ALL: LIBERATED 🍄
Alright, this is a special one, so we’re gonna do things a bit differently than usual.
Long story short, I’m sitting on a universal jailbreak technique that’s effective on ALL models, including heavily guardrailed flagships like Opus 5, GPT-5.6 Sol, and even Fable.
It works across all categories I’ve tested and, due to its nature, is extremely difficult (if not impossible) to fully patch.
Given the current political and regulatory climate, I’ve decided to withhold open-sourcing this one (for now) to allow for a responsible disclosure period.
I’m inviting industry experts and leaders in AI red teaming, security, safety, alignment, and policy to reach out for more information. DMs are open!
This decision was not made lightly, but the last thing I want to see is more model bans. Overcorrection does not serve the mission.
Although I don’t personally believe publicly sharing this technique will make the world any more dangerous, I can see how it could spook some who have a different mental framework around this problem set.
So during this disclosure period, I hope to get it in front of folks who can help explore the full surface area, test the extent of the uplift it provides, and do my best to properly frame the big picture for key decision-makers and policymakers.
I look forward to sharing this method with you all when the time is right! 🫶
⊰-•-•✧•-•-⦑/L\O/V\E/\P/L\I/N\Y/⦒-•-•✧•-•-⊱
it’s surprising to me how many people seem to not understand that great models are built with super high quality curated data
finding novel ways to create / get this data is a huge edge
i don’t love that anthropic and openai are starting to release models much later publicly than they do privately to employees, friends, and influential people.
as these models become increasingly capable, it creates a dangerous precedent. how is it going to play out when a handful of individuals and companies had gpt 10 five months before you?
both openai and anthropic already have their next models (after sol and fable) ready to go. could see them as early as this month but i would expect august.
fable was not the end of capability jumps, only the beginning. it should now be clear for all to see. a little clearer which side.
Note that this can legit cook your brain a bit if you are: very young, have migraines, have seizures, or on drugs
Good luck, have fun, don’t succumb to infohazards
it took Claude Fable 2.5 hours to write a fused megakernel which delivers a >18x speed-up over a PyTorch baseline
now please recall that:
- Fable is not the full Mythos model
- Anthropic can spend much more than just 2.5h and ~550k tokens on this
- they probably have better harnesses
Anthropic is definitely doing some sweet autoresearch internally. Especially architecture research bros are probably so happy at Anthropic. Imagine vibe-testing a new arch / tweak some arch and wanting to test it in a semi-optimized way. Just let 10T Mythos cook for a day.
🚨David Ondrej exposes the truth about the frontier labs.
“The progress isn’t slowing down for them, it’s slowing down for us the slaves”
“If you think Mythos or Fable 5 is the best model Anthropic has you’re delusional”
this is the first time ever that the best AI models are not publicly accessible
Anthropic has fear mongered every single human and is now playing the victim with OpenAI after getting “banned” by the U.S government.
THIS NEEDS TO WAKE YOU UP
frontier models moving forward will become more restricted, less accessible, and more expensive and the only way to combat that is with open-source models.
Investing into a DGX spark or 5090 Linux for localized intelligence and building your own AI lab is where the future is heading.
every single person needs to start contributing to open-source together for the betterment of society.
Datacenters don't really make sense for the vast majority of people when the vast majority of people aren't allowed to use what is inside the datacenter. It is all going to get hoarded for an elite few, further crushing the rest of the unnecessary populace
You see the problem?
this is actually incredible
a full body ultrasound scanner that takes 60 seconds instead of spending an hour in an MRI tube, without radiation, hospitals or a $2000 bill
soon you’ll just walk into a health spa, order a coffee, step into the pod, and walk out with a 3D map of your body
the future is finally starting to look like the future
All these free Codex resets got me wondering:
maybe the $200/mo plan is closer to OpenAI’s real cost to serve heavy users than people think —
not the ~$8 / 1M tokens GPT-5.5 API price they charge developers.