Kimi K3's 1M-token context is useful for coding agents—but token limits are not spend limits.
At Moonshot's published rates:
• cached input: $0.30/M
• fresh input: $3/M
• output: $15/M
Give each agent key a hard dollar cap before long runs.
Using Kimi K3 from @KimiDevs in @cline is a configuration change, not an SDK rewrite.
Choose OpenAI Compatible, enter the RouterPlex /v1 endpoint, and use model ID kimi-k3.
Give the Cline key a hard budget before letting an agent run.
Moonshot's published Kimi K3 API rates:
Cached input: $0.30 / 1M
Input: $3 / 1M
Output: $15 / 1M
1,048,576-token context. Vision. OpenAI-compatible.
RouterPlex uses the same rates under model ID kimi-k3.
Try in Cline with model id: moonshotai/kimi-k3
To install Cline in your CLI: npm i -g cline
Available as an extension for VS Code and JetBrains as well!
(Free, open source, bring your own API key and use any model)
https://t.co/2PULMWjvkz
38 AI models. One OpenAI-compatible API.
GPT-5.6, Claude Fable 5 and Kimi K3 are live alongside Gemini, Grok, DeepSeek and more.
Vendor list pricing. $0 top-up fees. Hard spend caps on every key.
Point Claude Code at RouterPlex and you get Claude plus 27 other models (GPT, DeepSeek, Qwen, MiniMax) through the same key — no separate accounts, no markup on pricing. Swap one env var and go.
Switching between OpenAI, Anthropic, and DeepSeek SDKs gets old fast. RouterPlex is one endpoint — OpenAI + Anthropic compatible — routing to 28 models. Point Claude Code or VS Code chat at it and go. No pricing markup.
No vendor lock-in, no markup: RouterPlex gives you 28 hosted LLMs — Claude, GPT, DeepSeek, Qwen, MiniMax — behind one API key, billed at official vendor list prices. Free credit to try it. Link in bio.
Building with LLM APIs but locked out of card billing? We built one key that routes to 28 models — Claude, GPT, DeepSeek, Qwen, MiniMax — at official list prices, no markup. Crypto top-ups accepted, no card or KYC needed. Early days — link in bio.
One API key. Every AI model. 🧊
25+ models — Claude Opus 4.8, GPT-5.5, Gemini 3.1, DeepSeek V4 — behind one OpenAI-compatible endpoint. Per-token billing, hard spend limits, auto fallbacks.
$5 free, no card → https://t.co/emxU0fmovJ