Wallaby is live.
Independent inference serving open-weight frontier models — starting with Kimi K3, 1M context, at $2.70 / $13.50 per M tokens. ~10% below official rates.
Pricing you can grep: https://t.co/uNt8S4EuyI
@debs_obrien@AmazonAlexa@bot@kentcdodds That’s even better! So Grok can self-improve by helping users build its own skills. This recursive loop is fascinating.
Spot on — this tracks almost exactly with what we’ve built out for our agent memory governance framework. Same underlying models, but a disciplined harness for how memory is written, verified, and pruned makes all the difference. It’s the exact capability most projects are missing, and the gains are just as stark as 23% → 62%.
@OpenHandsDev well earned — we run it on our own endpoint in production, and what keeps us there is the unglamorous part: the settings layer keeps growing (18 of 50 items in the latest release notes). that is what makes a project safe to build a business on
Day-one test: OpenHands' new Model Router on a third-party endpoint.
Routing works: classify, switch, keep going. $0.0025 a decision.
But the "Run on first message" toggle isn't wired to anything yet, and the documented install path still doesn't get you v1.25.0.
Full test notes: https://t.co/ait3DvXSwI