๐ Better inference efficiency, lower costs, broader access.
MiMo-V2.5 Series API pricing is now permanently reduced โ by up to 99% compared to previous pricing.
โจ Unified pricing across all context lengths.
MiMo Token Plans have also been upgraded:
โข 5โ8ร more usable tokens at the same price
โข Simpler and more transparent billing rules
๐ As a thank-you to current users, all current Token Plan credits will be fully reset.
๐ง MiMo-V2.5-TTS remains free for a limited time.
โฐ Effective May 26 at 6:00 PM PDT.
These improvements are powered by continued inference optimization and serving efficiency upgrades across the MiMo stack.
๐ ๏ธ Weโll also publish a detailed technical blog on the inference optimizations later โ stay tuned.
finally found the right metaphor for this shift in how i use opencode.
i used to treat it like 3D printing, where you build the thing layer by layer and commit to each piece as you go
now it feels more like progressive rendering, you start with a blurry version of the whole thing, then keep making full passes over it, and each pass sharpens the entire shape
doing this with gpt 5.5 and voice prompting is the first time things feel like they're clicking