We built Death by Diet Coke in less than 30 seconds using Cerebras Code Pro, where you get higher rate limits and more power for Qwen3-Coder.
We are opening the same number of Cerebras Code Pro/Max plans as diet coke cans in the office.
First come, first serve.
https://t.co/iwKPhjO5Ns
OpenAI GPT-OSS-120B is live on Cerebras
3,000 tokens/s - fastest OpenAI model on record
1 second reasoning time
131K context
Go build something fantastic.
https://t.co/jREGhLI2nj
Cerebras has been demonstrating its ability to host large MoEs at very high speeds this week, launching Qwen3 235B 2507 and Qwen3 Coder 480B endpoints at >1,500 output tokens/s
➤ @cerebras now offers endpoints for both Qwen3 235B 2507 Reasoning & Non-reasoning. Both models have 235B total parameters with 22B active.
➤ Qwen 3 235B 2507 Reasoning offers intelligence comparable to o4-mini (high) & DeepSeek R1 0528. The Non-reasoning variant offers intelligence comparable to Kimi K2 and well above GPT-4.1 and Llama 4 Maverick.
➤ Qwen3 Coder 480B has 480B total parameters with 35B active. This model is particularly strong for agentic coding and can be used in a variety of coding agent tools, including the Qwen3-Coder CLI.
Cerebras’ launches represent the first time this level of intelligence has been accessible at these output speeds and have the potential to unlock new use cases - like using a reasoning model for each step of an agent without having to wait minutes.
Cerebras Code: 20x faster than Claude, 1x the price
Today we are launching two monthly coding plans:
➡️Cerebras Code Pro: $50/m – for indie developers
➡️Cerebras Code Max: $200/m – for power users with 5x rate limits
Both plans get: Qwen3-Coder at 2,000 tokens/s, 131K context, and no weekly limits.
Sign up now: https://t.co/hzAROAbTqk
🟧🟨QWEN3 CODER is LIVE on Cerebras 🟨🟧
2,000 tokens/s - 20x faster than Sonnet
0.5s time-to-full-answer
131K context
$2 per M input/output tokens
Available in Cline, Windsurf & more
🟪 Qwen3-235B Thinking is live on Cerebras
1700 TPS • 131K Context • 1.7s time-to-full-answer
$0.6 | $1.2 per M tokens
Try chat: https://t.co/50vsHCl8LM
Get API key: https://t.co/JrvkzF6MS1
Pay-as-you-go via @huggingface
🟪 Qwen3-235B 2507 Instruct is live on Cerebras
1400 TPS • 131K Context • 230ms TTFT
$0.6 | $1.2 per M tokens
Try chat: https://t.co/39xaLQwNfj
Get API key: https://t.co/xIDW4v5BcU
Pay-as-you-go via @openrouter
Cerebras just beat NVIDIA Blackwell
Last week: Blackwell hit 1,000 t/s on Llama 4.
Today: Cerebras hit 2,500 t/s on the same model, same benchmarks by @ArtificialAnlys
Blackwell smoked Groq, AMD, Google – everyone.
Only Cerebras stands – and we smoked Blackwell.