Some of the worst acting I’ve ever seen in my life alongside the most predictable storylines and the most ridiculous styling. I’m obsessed I need 14 seasons. #AllsFair
OpenAI GPT-OSS-120B is live on Cerebras
3,000 tokens/s - fastest OpenAI model on record
1 second reasoning time
131K context
Go build something fantastic.
https://t.co/jREGhLI2nj
Cerebras x @IBM : Partnering to help enterprises accelerate AI adoption without compromise.
Businesses shouldn’t have to choose between bleeding-edge AI performance and enterprise-grade reliability.
It’s no longer enough for AI to be powerful, it also has to be practical, efficient, and trustworthy at enterprise scale.
That’s exactly why we’re working with IBM and @IBMwatsonx
Cerebras launched inference just 8 months ago.
Today it is officially part of Llama API.
Any developer can now click a button and get a wafer-scale chip to generate tokens at ~2,600 t/s.
Insane progress.
Cerebras and @Meta Collaborate to Drive Fast Inference for Developers in New Llama API
🦙The world’s most popular open-source models — now with the world’s fastest inference.
🔑 Native to @AIatMeta Llama API with 1-click API key generation.
🗣️ Unlock next-generation applications like real-time voice assistants, instant agents, sub-second reasoning
Llama 4 is here and it’s coming to Cerebras!
Starting next week, we will be serving Llama 4 on Cerebras Inference at instant speed. Thank you to the @AIatMeta team for your partnership.
Be the first to get access here: https://t.co/2SKkBZDBu2
🚨 Big news! Cerebras is launching six new AI data centers across North America & Europe.
This expands our capacity to 40M+ tokens/sec, making Cerebras the first hyperscale cloud for high-speed AI inference.
Read more: https://t.co/Bxk1QhGfzJ
I'm convinced India will never be able to compete with the US and China in technology if we keep treating it as a spectacle.
A bunch of influential folks are organising, Asia’s “largest AI event” in Mumbai later this month, and the speaker lineup has Bollywood celebrities, cricketers, and YouTube influencers. These are people who haven't written a single line of code in their lives.
A country doesn't become a technology leader through celebrity endorsements or political speeches. India will never be a technology powerhouse if we parade technology as an accessory. Real AI innovation doesn’t come from celebrity panels—it comes from builders. PhDs, engineers, founders—people who write code, build models, and deploy systems at scale.
The US and China didn’t lead in AI because of influencer summits. They did it through university labs, open-source contributions, and startup founders building from first principles. We need to build an ecosystem to listen and learn from builders.
Technology isn’t a spectator sport. If India wants to lead, we must put real builders at the center of the conversation.
DeepSeek R1 70B is now on Cerebras!
- Instant reasoning at 1,500 tokens/s – 57x faster than GPUs
- Higher model accuracy than GPT-4o and o1-mini
- Runs 100% on Cerebras US data centers
https://t.co/jREGhLI2nj
There are 4500+ NeurIPS papers... 🤯
The NeurIPS Navigator lets you search, summarize and instantly chat with the 4500+ papers accepted into NeurIPS 2024, powered by Llama3.1-70b on Cerebras.
👉 https://t.co/rvp9peKnVU
🚨 Cerebras Inference is now 3x faster:
Llama3.1-70B just broke 2,100 tokens/s
- 16x faster than the fastest GPU solution
- 8x faster than GPUs running Llama *3B*
- It's like the perf of a new hardware generation in a single software release
Available now at https://t.co/39xaLQwNfj