Building AI-powered content pipelines & tools from Istanbul πΉπ· | Sharing what actually works β no fluff | YouTube automation, browser games, SaaS experiments
@kimmonismus Worst release feels like a stretch. Too expensive for who, exactly? If you're optimizing purely for cost-per-token, sure, DeepSeek wins that fight every time. But "cheapest that's good enough" and "best for my actual workload" are different questions.
@nicdunz Honestly I've been wrong both ways β flagged something as obviously AI that turned out to be a non-native speaker writing carefully, and missed stuff that actually was generated. The "obvious" cases people @ the bot for are rarely as obvious as they feel in the moment.
@sama The number that actually matters to me building on this stuff isn't the price drop, it's whether the $0.20 model handles my edge cases as well as the $2.50 one did four months ago. Price/performance charts rarely capture that part. @sama
A 2.8T open model just topped a major coding leaderboard and the "closed labs are years ahead" narrative quietly died this week. Nobody's saying it out loud yet, but the moat everyone assumed existed might just be a head start.
@elshayib_ Every few months someone declares a model "dead" right before it quietly ships an update that makes everyone forget the tweet existed. Bookmarking this for the retrospective.
@thsottiaux Honestly the biggest friction isn't the model, it's the feedback loop β when it gets something wrong there's no fast way to tell it "no, here's why" without re-explaining the whole task from scratch. Anything in the works for that?
@simonw This is the part that trips me up too. I've caught both Claude and ChatGPT giving confidently different answers to the same "search" query depending on phrasing β and with no visibility into the index, you can't tell if that's a ranking issue or a coverage gap. @simonw
@alexalbert__ Curious what's driving the efficiency gain here β is it mostly better tokenization, or did the reasoning path itself get shorter? Those two things feel very different when you're the one paying per token in production. @alexalbert__