Claude Code's combination of Fable 5.1 and Opus 5 is significantly better than Astra/Sol/Terra. Grok isn't part of the conversation yet but I hope it can get there.
Opus 5 somehow went from being the laughingstock of the AI world to a phenomenal worker. Also the mileage you can get against weekly limits is remarkable.
Astra simply can't compete with Fable 5.1 on being the trusted, frontier-level agent I'm looking for. Astra may get a better score on the test, but Fable 5.1 feels like a real person who has a good feel for the world. It also has a backbone, whereas I can easily convince Astra to go against something it just recommended. I also had this impression with Fable vs. Sol, but that calculus was different because Sol was much more token efficient. Astra destroys the weekly limit.
Anthropic is sort of the villain right now but you can't deny how good its product is
As a former Opus 5 hater, I've come to appreciate its capabilities, especially given it's basically free with Fable eating all of my Claude Code weekly usage
Random thought on the frustration with coding benchmarks: the Michelin guide doesnβt over quantify a restaurant experience. Why do we do the same with these models?
Maybe the solution is to accept that just like so many things in this world, models canβt be reduced to pure numbers
Random thought on the frustration with coding benchmarks: the Michelin guide doesnβt over quantify a restaurant experience. Why do we do the same with these models?
Maybe the solution is to accept that just like so many things in this world, models canβt be reduced to pure numbers
@stefanoscalia Could not agree more
Camino Alto is an exception. They have a really good local rockfish dish. Arguably their signature entree at this point
I'd go as far as saying Astra is a regression for OpenAI.
I want an agent that I can trust with my highest level coding work. Fable 5.1 dominates in that regard. Astra can be unreliable, it often stops prematurely, and it destroys usage limits.
Fable 5.1 for planning and orchestration, and Sol/Terra for tier 2 tasks and below
everyone using fable/astra (most of you on subs, because no one can afford that) should try switching back to the other high reasoning models (opus/sol)
you'll likely realize your tasks dont perform any differently