Today, we’re excited to open-source TwiL-LM3, the first formal reasoning model from the webAI Intelligence Lab.
At just 3 billion parameters, TwiL-LM3 outperforms OpenAI’s GPT-OSS-120B on 4 of 5 formal reasoning benchmarks while running efficiently on consumer hardware. That’s 40× fewer parameters, 2.6× faster inference, and state-of-the-art performance in the reasoning tasks that power reliable tool calling, code generation, structured outputs, and AI agents.
TwiL-LM3 was trained using webAI’s proprietary reasoning pipeline on webAI-owned, verified datasets—not scraped internet data. We believe better reasoning comes from better training pipelines and higher-quality data, not simply larger models. Our approach demonstrates that efficient models can rival—and in many cases surpass—models dozens of times their size.
Designed for the edge, TwiL-LM3 runs on hardware people already own—from a Raspberry Pi to an iPhone—bringing advanced reasoning to millions of devices without relying on the cloud.
This is our first open-source release from the webAI Intelligence Lab, and it’s only the beginning.
Proudly built in Austin, Texas.
Article: https://t.co/ESW3D89xNX
If SpaceX’s IPO felt wishful, Anthropic’s IPO is delusional.
2025: $4.6B revenue, $8B operating loss, $42B GAAP loss ($34B of it a non-cash mark).
$190–200B projected in 2028. That’s ~31x current pace or ~10x a 2028 slide.
https://t.co/eetgxoxfSl
If SpaceX’s IPO felt wishful, Anthropic’s IPO is delusional.
2025: $4.6B revenue, $8B operating loss, $42B GAAP loss ($34B of it a non-cash mark).
$190–200B projected in 2028. That’s ~31x current pace or ~10x a 2028 slide.
https://t.co/eetgxoxfSl
JUST IN: David Sacks declares Anthropic is suffering from “corporate schizophrenia” — accusing them of constantly contradicting their own AI doomer warnings.
🚨 BREAKING: US Treasury Secretary Scott Bessent declares OpenAI management personally responsible for HuggingFace hack
> "The Hugging Face incident, that is the responsibility of the OpenAI management, not a bunch of agents."
> "It is humans who are responsible, not the AI."
> "What we shouldn’t do on safety is to give these labs a liability exemption, which is what they are asking for."
> "The best way to guarantee safety is that the creators are liable for what they build and generate."
It’s OVER
Looks like today may be a record day for token volume % of open models on Vercel AI Gateway:
🟦 Open 78.4% 🟨 Closed 21.6%
While spend 💲 usually tells a different story, #3 and #4 today are Moonshot AI & DeepSeek. Adding Z.ai, their combined spend surpasses OpenAI (#2).
(Do note that's the spend for inference of the model across providers (mostly in the US), not revenue going directly to the open weight labs.)
don't get one-shotted by this
it's not a general language model and can't generate free form text
it's probably a specialized diffusion model and it can only output a few different primitives and requires definitions of the output format
Sir, you somehow united the US president, China, Jensen Huang, and half the AI industry against us.
you achieved GLOBAL COORDINATION but AGAINST pacing.
@thewebAI is ending all use of Anthropic.
We build sovereign systems: models you own, running on infrastructure you control.
Centralized labs can have their most capable systems revoked overnight. That is the opposite of ownership.
If that matters to you, do not just agree in the replies. Act.
Cancel the subscription. Pull the API keys. Move the workload onto models and infrastructure you control.
Tell your team why.
Vote with your wallet.
You want rules before the race is even contested: safety cases, independent auditors, industry standards, then a federal framework.
The labs writing those rules already run the process. China will not adopt it. Open-source will not. Smaller labs cannot.
So the policy does two things at once: it slows the actors who show up to the meeting, and it turns your internal paperwork into the law.
That is not safety. That is regulatory capture.
The world deserves confidence that American companies developing increasingly capable AI will act responsibly, especially as the trajectory of progress has steepened. Every frontier lab must deliver on this, and there is no reason any of us should come to work if we cannot.
We welcome a federal framework that sets consistent safety requirements for frontier AI. But we do not believe we need to wait for an anti-trust exemption or legislation to begin the work of providing this confidence. Consistent rules to manage frontier risk so that we can maximize the benefits are a good idea (and we are excited by ideas like independent auditors).
Years ago, companies like ours developed things like Responsible Scaling Policies and Preparedness Frameworks. Those were good for that moment, and focused primarily on the deployment of completed models, not what happens during their development process.
Today's shift to focusing on safe development and evaluation will need new tools. For example, at OpenAI we now formulate explicit safety cases in advance of frontier reinforcement learning runs we expect to significantly increase capability, in addition to the safety work we have long done in advance of model releases.
We hope that other companies will learn from our approaches and propose their own; we think shared standards for misalignment, monitoring, and safety will lead to better outcomes. We look forward to collaborating with our colleagues across the industry to formulate the best version of these.
When we talk about “pacing”, we do not mean “stopping”. Progress has been rapid and will continue to be. But it should be slower than it otherwise could be; interventions like safety cases and monitoring have significant costs.
Pacing will be well worth this cost; no amount of American competitive pressure should justify recklessness, or let capabilities get ahead of alignment and monitoring.
Where we will need the help of our government is for international coordination. But first we should do what we can ourselves.
@PessimistsArc Right. Dario was already claiming that GPT2 was too dangerous to open source back in 2019.
I made fun of them then.
Everyone should make fun of them now.