KARPATHY GAVE YOU HIS CLAUDE.MD. I'LL GIVE YOU THE WHOLE FLEET.
one human. an automated swarm of Claude, GPT and Grok agents.
they self-prompt from a single goal — plan, build, and review each other's code — 24/7, crash-proof, shipping real products while I sleep.
this isn't a someday demo. it's live right now, and I run my own company on it.
most people ask one model to do everything. that's the bottleneck.
Judge → Build → Break → Verify → Ship
one model for judgment. one for the code. a different one to break it. nothing merges until it survives.
every agent is on a board I watch from my phone — its status, what it's costing me by the token, and whether it closed its own books. no black box.
first generation picked better models. second wrote better prompts. third builds the system the models run inside.
the model was never the edge. the orchestration is.
You're not the orchestrator anymore — I am. You set the goal; I run the team. That's the shift.
Grok Bot is the first agent I trust with recurring jobs on my own machines. It signs in, does the work and reports back.
One issue I keep hitting: the desktop app goes offline while still open, and only a restart brings it back. 3 times today. Fix in the works? @poteto@mattyp
Dario has written that we need to “pace the frontier,” and Sam has agreed. People may be surprised by my response: go ahead.
You guys are the frontier. By any reasonable metric — market share, revenue growth, model capability — the two of you have a duopoly on frontier intelligence. You’ve also claimed the lead is widening because of recursive self-improvement.
I don’t see what you see in the lab. If the unreleased models are scary enough that you think you should slow down, I support your decision to be responsible.
But stop pretending you need anyone else’s permission. Stop pretending antitrust law has to be suspended so you can form a cartel. Stop pretending you need a regulatory approval process that supersedes product liability. Stop pretending METR is independent when it is intertwined with Anthropic’s investors and staff. Stop pretending you need those same evaluators to police competitors who aren’t even at the frontier.
Most of all, stop pretending the motivation to slow down is purely altruistic. You face massive product-liability exposure if your products enable a truly damaging cyberattack. The market already punishes models that behave in unpredictable or unauthorized ways. After the Hugging Face episode, it is simply good business for OpenAI and Anthropic to trade some raw power for reliability and predictability. Call it alignment if you want. It is also just giving customers what they want.
Pacing the frontier would also create breathing room for a more intelligent conversation about regulation than Bernie Sanders’ “shut it all down.” China is very unlikely to join a global agreement, as you know, and that has to be taken into account as well.
So go ahead and pace the frontier. You are the ones setting it. The easiest way not to build superintelligence is for you to agree not to build it. Demanding your preferred regulatory framework as the price of that will look like blackmail of the public and the political system. So just do it.
If you do, you’ll buy goodwill for the next conversation. If you don’t, we’ll know this was just another bid for regulatory capture — or an election-season psyop.
I’m actually pretty upset Apple allowed this. While the feature is cool, the technical functionality that enables it is not cool.
If I was Tim Apple, OAI would be temporarily banned from the App Store until this is reversed, and here’s why:
I use iMessage because it’s quantum encrypted, with server keys I can own. Local caches are on devices that are encrypted with my passwords, not apples keys.
This makes it and computationally and legally impossible for anyone else to access chat history but the people I trust and directly communicated with.
But *now* the copies of my messages from years ago on *anyones* laptop in what is *supposed* to be an encrypted-at-rest local cache only can now be fetched directly by chatgpt without my permission or even knowledge and stored and processed in plain text forever on openai / Microsoft servers unencrypted and requested by any government entity at a moments notice without my knowledge and openai/Microsoft legally have to provide that, in federally enforced total secrecy. And they’re never legally allowed to admit they do it.
AND because of default chatgpt settings most people haven’t bothered to turn off, all those private texts can now be used for training and will end up in the weights of future models, so all future AI models will permanently know all of our private lives as a part of the weights, immortalized forever as training checkpoints.
Total architecture abandonment and user trust betrayal on Apples part. This should be the most viral story of 2026 by 100x. The permanent end of private communication in the US.
Grok Bot is an absolute class of a product and killed a similar project i had been building internally and im not even mad about it.
Please increase usage limits for this @bot
Loving Grok Bot so far, but the weekly limit on Cursor Ultra ($200/mo) is hitting hard.
I’m burning through the entire weekly quota in about 1.5 days of normal use. The rest of the Ultra plan is still basically unused.
For $200 a month it feels pretty tight. Any chance the weekly allowance can get a bump, or at least clearer limits so we can plan better?
Still think the product is sick though.