@kimmonismus Hey Chubs,
Anthropic doesn't want you to use Opus as part of a workflow you can depend on. That's what they keep saying.
I heard Dario made them pivot hard toward creating a $200 per month Mystery Box subscription.
@thsottiaux@aarondotdev Tbf using AI does not magically generate well thought out code. And tbf, using human developers does not always generate well thought out code.
Strawman!
Who cares. Only the US labs and their grifters. Until ox-alpha, people dealt with dario and sam fatique because their product was better. Their lead is shrinking at an exponential rate. Its clear now that money cannot buy 1st place for much longer. The US labs need better leadership.
@MiaAI_lab Running it inside OMP with an Erdos harness. It is genuinely refreshing. It stopped the validation theater, reframed the problem and finished what my Anthropic and OpenAI clankers couldn't.
https://t.co/JqHS5HUCt9
Ox Alpha is frontier. wtf. Seriously wtf. Tencent? Xiaomi? Really? It's more well-done than any other Chinese model. I swear, it's better than Kimi K3. It's good *in ways that Chinese models are typically horrible at*. It's FAST
I can half-believe it's a stolen Claude checkpoint.
1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because it really brings out the heart of an important conversation.
First, on regulation, I think that “either concentrate it in the hands of a chosen few companies and politicians via regulation or distribute it widely” is a false choice. I know that there’s a sort of Silicon Valley shorthand where regulation = regulatory capture = concentration of power, but I’ve always found this to be an overly simplified picture of the world. Many people outside this bubble think of regulation as something that constrains corporate power and benefits ordinary people. I don’t necessarily agree with that perspective either, rather I think it’s complicated and really depends on what the “regulation” consists of. But in particular I think that those in the “regulation = regulatory capture = concentration of power” frame often underrate the decentralizing power of objective and fair institutional processes. A crude analogy is that the formal court system can sometimes feel stuffy and elitist, but it does a much better job of defending the rights of vulnerable individuals than the alternative, mob justice. At their best, institutions can vest power in ideas rather than people, and thereby decentralize that power.
This is why Anthropic has always made its policy proposals very carefully. We try very hard to make proposals that disadvantage (slow down) frontier AI companies while *advantaging* smaller competitors. California’s SB53 (which we supported), and even the much-maligned SB 1047 (which we were ambivalent on), completely exempt any company below a certain amount of revenue or model training costs from being covered at all (it was $500M for SB 53, lower for 1047 but we objected to that). More recently the testing process we’ve advocated for at CAISI and the White House involves more rigorous tests for frontier models than off-frontier models — something that differentially advantages challengers. Similarly, the “Pacing the Frontier” letter envisions (or at least Anthropic’s preferred implementation of it envisions) modulating the pace of the very best models while not constraining those who are catching up. This hurts the business interests of the frontier labs and helps challengers, including open-weights!
Overall my view is that AI is *structurally* a technology that tends to concentrate power, for reasons that have nothing to do with regulation (more to do with the extreme implications of the scaling laws). Open-weights do help some with this but are nowhere near a sufficient solution because they simply shift the concentration somewhat to those with the most compute and chips (which are roughly the frontier labs plus maybe hardware providers). By contrast I think the right “rules of the road” can simultaneously (a) address AI’s cyber/bio/alignment risks, (b) institutionally constrain the power of the frontier AI companies, and (c) leave room for open-weights models while also addressing the specific risks that they bring.
BTW I do not think that the events of the last few months have “failed to result in [my] preferred regulatory path”. The approach that the Trump administration is reported to be taking — pre-deployment testing for frontier models, and also testing of open-weights models when they get closer to the frontier — is one that I am very supportive of, though of course I have to see the details to be sure. I am also supportive of Demis Hassabis’ ideas around a FINRA-like entity. This contrasts with six months ago when most of the industry was still pushing for preemption of all state regulation and no apparent federal approach either.
@Agusx1211 Remember when the first major problems started shortly after the Opus 4.6 golden era? Anthropic didn't even realize there was a problem.
have you tried this ->
https://t.co/aIDt4ESvs2