Cheap inference is usually slow. Ours isn't.
GLM-5.2 on Merius: top-tier throughput, bottom-tier price.
$10 free credits for every new signup โ limited time
https://t.co/HmgJgjGI6M
Kimi K3, now flat-rate in Merius Frontier ($149/mo) ๐
No metering, no waitlist โ point opencode, Cline, or your agent at it and go.
Free 30-min test available on our dashboard.
(Also on per-token API if you prefer โ $2.50 in / $13.00 out per 1M.)
@TencentHunyuan Love this. The 1-bit/4-bit release makes Hy3 incredibly accessible for local runs.
For teams who'd rather not manage the GPU, we put Hy3 on a hosted API at Merius โ free all of July. Same model, zero setup: https://t.co/yhz7LnARLZ
Hy3 is free on Merius for all of July. โก
Tencent's 295B MoE โ reasoning, coding, agents โ served fast, at zero cost this month.
Swap your base URL, keep your OpenAI SDK, start building.
https://t.co/vgki0LY6Ci
@TotalWorldApps Yeah GLM-5's solid. Scaling usually comes down to the serving layer more than the model โ that's the part we handle, so throughput holds up as you scale.
Cheap inference is usually slow. Ours isn't.
GLM-5.2 on Merius: top-tier throughput, bottom-tier price.
$10 free credits for every new signup โ limited time
https://t.co/HmgJgjGI6M