If your team is using open-source models, there’s a good chance Auriko can reduce inference costs by 10-40%.
We benchmarked 80K+ API requests and saw ~30% average savings from Auriko's cache-aware routing.
Read the statistical report here: https://t.co/EH0N7ATQmF
Auriko won #1 Product of the Day on Product Hunt! Huge thank you to everyone who supported us.
We’re building Auriko to make LLM inference cheaper, easier, and more reliable for AI developers.
This means a lot to us. Thank you!
Your AI agent can now manage its own inference infrastructure.
We shipped the Auriko Management API.
Agents can rotate API keys, check budgets, swap provider credentials, query usage, and create scoped child keys for sub-agents.
This is built with production security controls: granular permission scoping, MFA-gated credential creation, and a full audit trail for every operation.
Your agentic system can now automate inference operations end to end.
For years, the standard startup advice was:
Build the core feature. Ship something scrappy. Validate fast. Clean it up later.
As we kept building, we've realized this premise is no longer universally true. In some case it’s counterproductive.
By popular request. We just shipped Responses API support on Auriko!!
Build stateful, tool-using agents with the newer agent API. Run response API across Claude, Gemini, DeepSeek, OpenAI, Qwen, Llama, Grok, MiniMax, and 250+ models.
One agent interface. Every model.
https://t.co/rIVdHfz7fv
By popular request. We just shipped Responses API support on Auriko!!
Build stateful, tool-using agents with the newer agent API. Run response API across Claude, Gemini, DeepSeek, OpenAI, Qwen, Llama, Grok, MiniMax, and 250+ models.
One agent interface. Every model.
https://t.co/rIVdHfz7fv
Silicon Flow is now live in Auriko!
Try Hy, step fun, and many new models in Auriko.
Working toward more model choice, better uptime, and lower cost.
@SiliconFlowAI