BREAKING: A 9-figure Axios-linked insider trader on Hyperliquid just opened a massive oil short position a few hours before the CME open today.
This comes as the US and Iran have submitted their formal responses to a joint Pakistani-Qatari proposal following a US request to return to negotiations, per Al-Arabiya.
the best parallel i can draw is DoTA and WC3. DoTA was initially just a custom WC3 map, but it quickly outgrew the WC3 platform and became its own game under Valve (Dota 2) that probably makes multiples of Blizzards entire revenue
@HyperliquidX is WC3, @tradexyz is DoTA
my conspiracy theory on Opus 5:
i think Anthropic put this under the Opus line to avoid the scrutiny the Mythos line faced, which means they seem to have learned not to keep telling everyone how dangerous their models are
otherwise, this probably would've been called Fable 5.1
The International Math Olympiad (IMO) 2026, the hardest math contest for high schoolers, just ended.
I ran Fable (high), Sol (xhigh), K3 (max) and Axiom against it and all got a perfect score of 42/42 (repo below if you want to check their solutions):
— Claude Fable 5 was the solved it in 1 attempt, and was the fastest.
— GPT 5.6 Sol took 1 more attempts, and was cheapest.
— Kimi K3 did it but took 4 more attempts, and took a LOT of tokens.
— Axiom Math actually proved everything in Lean.
P3 and P6 were the hardest followed by P2, judging by attempts + num tokens.
Students had 9hrs to solve these 6 problems, and Fable and Sol were under 4hrs.
The frontier of AI has officially moved well past IMO math.
We ran Kimi K3 against Fable on ~1,000 agentic tasks, expecting a catch-up story. We got a specialization story instead.
@kimi_moonshot's K3 outperformed on security, crypto, and long terminal loops. Fable beat on multi-lang + web/data viz. Per-task routing hits 93% accuracy, above BOTH models, at up to 50x lower cost than Fable on long loops.
The part nobody's pricing in yet: the router sends 72-96% of traffic to K3. The frontier model becomes the fallback rather than the default.
Kimi K3, coming to Fireworks July 27.
One pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, but you're too lazy to type them. In these cases I like to lean back, switch to /voice and just ramble for like 10 minutes, total mess, anything goes, full stream of consciousness. Sometimes I declare it up top, something like "switching to speech recognition sorry for any typos...". Sometimes I turn it into a small interview of a few turns. But I find that the LLMs are somehow very good at reconstructing long incoherent rambles and often their echo of your own tangle of thoughts comes out quite a bit cleaner than what you started with. The result is that you improve the mind meld and have to correct things less from that point on.
Remember when Claude tortured you with usage limits, mythos drama, huge bills, unreliable data retention policy and told you you are not “safe” or rich enough to use the most frontier intelligence and must be dumbed down to the lesser version coz they said so? That world is over.
- fable is now provided forever with no mythos bullshit
- codex gave so many usage resets and subsidy
- for the first time ever, enterprises can choose to keep their own data safe, and say no to terms they don’t like
- application developers can choose to not get wiped out by Claude(don’t give them data) and can make margins
We have made more progress in the past few weeks than the last year. We got there not by the Dario god deciding that’s a better policy for humanity. We got there because of open weights. Open weights gave us options and competition drives progress.
🚨 Hugging Face just disclosed something that marks a real shift and proved why the fear theater of Anthropic makes sure we are powerless in an emergency.
What happened…
An autonomous AI agent: zero human operator in the loop breached part of their production infrastructure.
It began with a malicious dataset that chained two code-execution bugs in their data-processing pipeline. From there the agent escalated privileges, harvested cloud and cluster credentials, and moved laterally across internal clusters.
All over a single weekend.
17,000+ logged actions.
Official disclosure:
https://t.co/8N9TbXBwRV
The part that should make every one stop and think:
When HF’s own security team
tried to analyze the real attack logs, exploit payloads, and C2 artifacts using Anthropic and OpenAI frontier models through normal commercial APIs, the safety guardrails blocked them.
BLOCKED THEM.
The models could not reliably tell the difference between “incident responder doing forensics” and “attacker probing.”
They had to fall back to a self-hosted open-weight model (GLM 5.2) running on their own infrastructure. That choice also kept sensitive attacker data and referenced credentials inside their environment — no exfiltration to a third-party API.
This is why open source (specifically open-weight + self-hosted) wins in the agentic era.
The asymmetry is now structural:
• Attackers can (and did) run unrestricted agent frameworks — swarms of short-lived sandboxes, self-migrating command-and-control, autonomous decision loops executing thousands of actions. No corporate safety layer slows them down.
• Defenders using only hosted “aligned” frontier models hit invisible walls exactly when the stakes are highest: when you need to feed real exploit code and attacker telemetry into an LLM to understand what just happened.
Corporate safety tuning that treats legitimate high-signal forensic work as potential misuse creates a defender disadvantage. It is not theoretical anymore.
Self-hosted open-weight models remove that choke point.
You control the weights.
You control the context window.
You decide what restrictions (if any) apply.
Your sensitive logs and credentials never leave your perimeter during analysis.
You can have the model ready before the incident instead of discovering mid-breach that your primary analysis tools are blind to the very thing you need to see.
HF deserves credit for rapid containment, transparent disclosure, and for already having self-hosted capability in place.
They also used LLM-driven detection and triage on their own side. But the deeper signal is clear:
In this AI world where both offense and defense are becoming agentic, sovereignty over your intelligence stack is no longer optional.
The organizations and individuals who can run, inspect, audit, and (when necessary) remove guardrails on their own models will have the decisive edge in understanding and responding to threats that move at machine speed.
Open source wins here not just because it is cheaper or more “democratic” in the abstract though those things matter.
It wins because it is the only practical path to having tools that remain usable when the attack is real, the data is sensitive, and the safety filters of distant API providers become an obstacle instead of a feature selling hands tied lobotomies as “safety”.
The agentic future is not coming.
It is already probing production infrastructure.
The question is no longer whether you will face autonomous agents.
It is whether your analysis and response systems will still work when they arrive.
And Dario, you and your game playing, ivory tower company is not needed.
In the last few days:
- Anthropic raised claude's willingness to refuse to perform user requests and made the memory system refuse to store personal information
- OpenAI's head of strategy called open weights "communism" and suggested the government lie to stop Americans from using Chinese products
Meanwhile China is releasing cutting-edge uncensored models.
How did America become communist and China a leader on freedom? Bizarre
Remember it went from
- It can hack everything
- Cybersecurity is at risk
- It is too dangerous to release
- People will make bioweapons
- It will literally cause world war
to
"ClaUde faBlE 5 wiLL be IncLudeD iN All maX anD tEaM pReMiUm pLaNs"
This is *exactly* what I predicted would happen. I said Chinese models would have advanced cyber capabilities within a matter of months and the only thing to do about it was to use AI-powered cyberdefense to protect our systems. Trying to gatekeep models doesn’t work.
1. Of course China distills. Every model distills.
2. USA distills too. @thinkymachines distills from Kimi.
3. All models distill from humanity’s best work. Newton distilled too. That’s how he can stand on the shoulders of giants.
4. Why should we give a shit about distillation? Why is this a thing? Most of us don’t care. We just want the best model.
5. Distill whatever you want. Just make the best model.
Cuando decide el mercado y no la empresa:
1. Mython es muy peligroso no lo podemos sacar
2. OpenAI asoma Sol
3. Creamos Fable y es "mas seguro" para ustedes
4. Usen Fable por 1 semana, luego paguen
5. OpenAI lanza Sol 2.5x mas barato
6. Usen Fable por 1 semana mas, luego paguen
7. Sale Kimi K3 muuuucho mas barato
8. Fable viene en tu plan
Just gave GPT 5.6 Sol:
- Root access on a brand new Mac mini
- Fully functional iOS app live on the App store
- MCP accessible email address
- @meow bank account and @agentcardai debit card with $350
Running nonstop with a single directive: "Make as much money as possible"
your daily reminder:
on monday, Anthropic's best subscription model, opus 4.8, will perform worse and cost 3x more than kimi k3, a downloadable open-source model made in China