@thsottiaux More reliable German STT. Idk what's happening but the quality varies so much. Usually, mobile worked great and within codex it was horrible. Last two days mobile is soo bad. This is really such a convenience driver for me. One of the reasons why I use chat over Claude
@thsottiaux The fact that people are scheduling runs in order to open the 5h window overnight so that they can start the day with 2.5h and 99% quota left tells you everything π
I have a one-test benchmark where I ask any new model to identify tennis courts from a Google Maps screenshot and calculate how many Padel courts could be placed there. No model so far was able to pass. GPT 5.6 Luna (!) x-high was the first model to crack it. π#codex#gpt
@tweets2xs@MTSlive Distillation is like taking a photo of an artwork, printing it an reselling it. It's not quite as good as the original, but it'll work for some people.
Alibaba supposedly used their access to Claude to train their own model based on Claude's answers.
I am finally a token billionaire - entirely without token maxxing. Also, thanks to #Fable, today was my highest API-price-equivalent day at $70. I know, beginner numbers for some, but I feel I'm on the right path!
Had Opus 4.8 develop a feature, then Fable 5 and Gpt5.5 code review. Passed the findings back to Opus and asked which reviewer seemed more senior. It went with 5.5.
#fable#gpt55#opus#codex
@theo@themmyleke Which is actually wild to me. I mean Google owns the vertical stack - chips, infra, model, application. Their unit economics have to be great. Plus, they have a money printer. Why they didn't position themselves as the cheapest option is so weird to me.
@thsottiaux@skirano@OpenAI handing out 10x extra usage for special projects.
My expectation: tools that speed up cancer research, environmental preservation simulation,...
Reality: an official chatGPT plugin that extends the Codex experience and will likely be swallowed within the month.
@depak_7@chrisgpt It's that far apart on the hardest problem set. Is all your dev work super hard engineering problem? If not, that explains why. Models are likely similar in performance for easy-medium problems.
@thsottiaux yooo you guys pulling 5.3-codex from the subscription is not cool. I've been using it for less complex work. Don't get why we're forced to use overpowered models for simple tasks. I guess I'll get a $6 MiMo subscription to swap in for the easier stuff.