This is concerning. For the first time, a Chinese model Kimi K3 has taken #1 on the Frontend Code Arena and is scoring at or near the frontier on other benchmarks.
Meanwhile America is tying itself in knots: politicians and bureaucrats are banning new data centers, piling on state regulations, and pushing for new federal agencies to pre-approve frontier models.
This is how you lose the AI race. The rest of the world won’t play by our rules if we bog ourselves down. Permissionless innovation is how America won the internet and became the technological envy of the world. We can do it again with AI -- while addressing risks in a targeted way -- or we’ll watch our lead evaporate.
Taiwan solved tax evasion in 1951 with a trick so cheap it should embarrass every tax authority on the planet.
The problem was an all-cash economy full of small shops. A merchant pockets the cash, skips the receipt, and the sale never existed. Auditors can't catch what was never recorded, and hiring enough of them to watch every noodle stand costs more than the missing tax.
So finance chief Ren Xianqun flipped the incentive. Print a lottery number on every receipt. Draw winners every two months on live TV. Top prize today: NT$10 million, about $310K.
Suddenly the customer and the shopkeeper want opposite things. The merchant wants the sale off the books. The customer wants the ticket. And there are millions more customers than merchants. Every transaction now carries a built-in witness demanding the paper trail.
Year one, reported tax revenue jumped 75%, from NT$29 million to NT$51 million. Seventy-five years later, roughly 70% of Taiwanese still play. Convenience stores redeem the smallest NT$200 prizes at the register, so even a coffee receipt feels like a scratch card.
The elegant part is what the audit force costs. The prize pool runs about NT$7 billion a year, roughly $20 million. In exchange, the government gets 23 million unpaid auditors working every checkout line in the country, forever. No inspector general on earth delivers that coverage at that price.
Greece, Italy, Portugal, and Slovakia all copied it. The most effective compliance tool ever built looks like a game, and that's exactly why it works.
I’ve had a number of conversations with folks inside and outside government about the current situation with Anthropic, and here is what I believe to be true:
— As we know, Anthropic publicly released its Mythos class models earlier this week under the commercial name Fable.
— Fable is Mythos with guardrails. But if those guardrails fail, then you’ve exposed Mythos and its advanced cyber capabilities to people who shouldn’t have them. (Keep in mind that Anthropic itself widely promoted the idea that Mythos was a cyberweapon and needed to be regulated as such. They asked for government regulation of Mythos and championed the guardrails on Fable. If there is a vulnerability — big or small — it is Anthropic’s responsibility to patch.)
— A highly credible trusted partner of both Anthropic and the USG who was testing Fable came forward with a jailbreak of those guardrails. The Admin asked Dario to fix the jailbreak or de-deploy the model. Dario refused.
— In their blog post, Anthropic defended its decision by saying the jailbreak isn’t serious. That is not what the trusted partner and the USG believe; nor is that kind of minimizing language consistent with Anthropic’s brand as the AI safety company. It’s difficult to fathom how they could claim a jailbreak allowing operability of a cyber weapon could be defined as not “serious.”
— In the past, Anthropic has always said that safety must be top priority and taken super seriously. In this case, Anthropic prioritized the continued offering of the consumer model over safety.
— In reaction, the Admin issued the export control. The Admin did this reluctantly. It’s been very surprised that Anthropic hasn’t wanted to cooperate with a reasonable safety request (ie fixing the jailbreak issue). Anthropic’s reaction is very much at odds with their branding and ethos as a safe AI research community.
— The Admin’s hope now is that Anthropic remediates the safety issue, the export control is lifted, and Fable goes back into general release. The Admin wants all of this to happen as soon as possible. It is frankly bewildered that Anthropic hasn’t wanted to comply with safety requests that it previously said were its highest priority.
— Those trying to misdirect and tie this action to the prior DoW/Anthropic issues are wrong. The Admin values Anthropic’s technical capabilities and feels that this issue, while serious, should be easily resolved. The ball is in Anthropic’s court.