At @linear we optimize for solving customers problems & craft – not A/B tests, internal frameworks/OKRs.
To optimize for craft, give the builders control. To solve customer problems well, try to make the whole team have customer understanding.
Then combine these for the magic
This "they distilled our models!" is starting to resemble the moment when Redis was winning in the database space really hard, and certain actors started with "but it is not linearizable!", and also attacked me personally. Eventually, a few started selling Redis, before closing.
few weeks ago, Fable 5 was so advanced it needed the gov and had to be taken offline.
today, we have access to arguably better models for $20/month.
took what? six weeks?
What's incredibly scary about lists like this is it shows how financially entangled the entire tech industry is.
This creates a cabal that enriches themselves, instead of competition that enriches the economy. Sad to see so many smart people participate.
@timwangyc@seedtosunflower@Joshuabrowder When do you expect to be out of early access ? Can't wait to get my hands on this to see what it can do :) currently a davinci resolve user but have literally been waiting/imagining something like this.
DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST time I feel I have a frontier model running on my computer. This is *crazy*, and probably a much stronger change in the landscape than PRO.
@jimkxa@tenstorrent You should just sell consumer/pro-sumer hardware that gives decent tok/s for these open but a tad larger models, eg... 300B + , would sell like hotcakes. Maybe this is the plan?
@Grummz I think games have a sort of lifetime and prime time, then people move on. Generational shifts, new things on the horizon. Even with WoW this happened, although it lasted a long while. Culture define the experience, when culture shifts so does the experience.
@scaling01 These puzzles are very simple, and it just goes to show current ML paradigms does not generalize very well at all. Purely in distribution interpolation. This is a good benchmark.
Maybe its true, @elonmusk should just add the 3d mmorpg component and let us roam the online metaverse using our x avatars/personas and move about in this 3d world - I envison it being cyberpunk like!
#apple ... why don't they buy something like Taalas and then buy Anthropic, encode #opus 4.6 on next gen mac silicon hardware, 16k tokens per second smth... people have an incentive to buy new macs every time a better model / chip comes out... thoughts @TheAhmadOsman
Five years out, when billions of coding agents exist, software development will be largely solved. The traditional moat of “we have better engineers” disappears. Products will be copied, improved, and open-sourced almost instantly. Defensibility will shift away from code itself, toward distribution, data, brand, and community.
@NeverSinkDev ~20 years as professional dev, here is my take. We wont work the same way in the future, but also my gut feeling is we will work. The AI has no desires, no emotions, no taste, no passion. This is what we contribute, we are the consumers of the output (eg s game). We are creators.
This makes me sad and angry, because claude code is a good product... come on, you train on the entire internet ffs, wtf is this? So you can ignore copyright, train on all publicly available code and text and images whatever, and then boo-hoo someone trains on your output.
We’ve identified industrial-scale distillation attacks on our models by DeepSeek, Moonshot AI, and MiniMax.
These labs created over 24,000 fraudulent accounts and generated over 16 million exchanges with Claude, extracting its capabilities to train and improve their own models.
Can someone explain to me how this model can be getting 77% on ARC-AGI-2 and still lose at Connect 4? The methodology says this is on the "semi-private" (not shared openly online) set which I assume is public to Google. Overfit? Lack of third party run evals? Connect 4 = AGI?
The Claude C Compiler is the first AI-generated compiler that builds complex C code, built by @AnthropicAI. Reactions ranged from dismissal as "AI nonsense" to "SW is over": both takes miss the point.
As a compiler🐉 expert and experienced SW leader, I see a lot to learn: 👇
Well damn, it was bound to happen and this morning it happened.
There's a big chunk of code touching many pieces that i know in depth because i proudly hand-crafted it all. But it has one bug that i wasn't able to pinpoint even after an hour of debugging yesterday evening.
This morning i resume debugging, but after another 10min of being none the wiser i decided to shoot a prompt to Opus4.6 and let it search while i continue debugging.
A mere 1min51s later, Opus actually found the bug, in a file that i didn't even consider looking at during my two debugging sessions. This marks the first time a LLM found the bug in my code faster than me. I've tried this many times in the past, and i was always faster or ~same.
FWIW the root cause was simple: https://t.co/u4h3qzRLxg(tuple_of_pyints) returns a np.int64, not a python int.
Finding this as the root cause of my bug is what was not simple: i didn't even mention that code part in my prompt because i didn't consider it.
Is it weird that AI coding assistance is not giving me identity fracture?
A lot of software developers are feeling disoriented and threatened these days. Programming by hand is clearly going the way of the buggy whip and the hand-cranked auger. Which is how we're finding out that a lot of people have their identities bound up in being good at hand-coding and how it feels to do that.
That's not me. It's not me at all. Rather to my surprise, I don't miss coding by hand, not any more than I missed writing assembler when compilers ate the world and made that unnecessary. (That was in a couple years back around 1983, for you youngsters.)
Maybe the fact that I'm not feeling any of this disorientation disqualifies me from having anything to say to people who are. On the other hand...if you can learn to emulate my mental stance and be completely unbothered, maybe that would be a good thing?
So. If you're a programmer, and you're feeling disoriented, try this on for size:
I like being a wizard. I like being able to speak spells, to weave complex patterns of logic that make things happen in the world. Writing code is a way to manifest my will.
Yes, I've piled up a lot of arcane knowledge over the 50 years I've been doing this. But languages of invocation, they come and they go. Been a long time since I've had any use for being able to program in 8086 assembler, and that's okay. I have better spells now, and these days some rather powerful familiars.
What I'm inviting you to do is think of yourself as a wizard. Not as a person who writes code, but as a person who is good at assuming the kind of mental states required to bend reality with the application of spells.
And if that's who you are, does it matter if the spells are painstakingly scribed in runes of power, versus being spoken to an obedient machine spirit?
It's all one; it's all the manifestation of will. Arcane languages come and go, machine spirits appear and then diminish to be replaced by more powerful ones, but you? You are the magic-wielder. Without you, none of it happens.
Same as it ever was. Same is it ever was. And so mote it be.
@signulll Extremely good, most software devs that I worked with have been pretty pessimistic in general, they wont probably get anywhere. The dooers, the builders, the optimists - will thrive
I’ve been thinking lately about what does make a company a startup.
Best I was able to come up with was a "star wars" test:
Does the company feel more like rebels or more like the empire?