For Local AI I'm using GLM 5.3 Flash (on 2x DGX Sparks) + Qwen 3.8 27B (RTX 5090).
This gives me a very nice combination of Intelligence + Speed.
GLM 5.3 Flash can spawn sub-agents that are very fast and control them.
GLM 5.3 Flash usually gives me around 30 tok/s and can spawn sub-agents using Qwen 3.8 27B with around 130 tok/s each.
Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family.
It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.
@MrSage Yeah, it's just the 3d files, I uploaded them to https://t.co/v7HDZLsnZ7 then tried to find the cheapest price in the world 😂 for me it's around 36 euros. Which sounds pretty expensive, given this is something very small, but I'm not going to buy a 3d printer only for that haha