🚀Claude Opus 5 from @AnthropicAI just arrived on Amazon Bedrock https://t.co/vqwPshniji
For teams building agents, Claude Opus 5 powers long-running agents that work for hours without oversight, recovers from errors, and pushes back on flawed instructions rather than executing blindly. It matches Fable 5’s top-tier intelligence on many domains.
It will be available on @kirodotdev as well.
Try it and let us know what you think.
Claude Opus 5, Anthropic's most capable Opus model, is now available on Amazon Bedrock & Claude Platform on AWS. 🚀 https://t.co/Noh8hJ3dWS.
Opus 5 delivers improved agentic coding, knowledge work, visual understanding, & long-running tasks. Zero data retention compatible for enterprise workloads.
Opus 5 is available today on all paid plans and the Claude API, priced the same as Opus 4.8. It’s the default model on Claude Max, and the strongest on Claude Pro. It’s also offered in Fast mode, which runs around 2.5× the default speed.
Read more: https://t.co/YMw7gLMT7e
Today, we're introducing GLM-5.2 Fast: our GLM-5.2 Model API designed for the most demanding real-time use cases.
The Fast tier delivers 2-3x higher TPS than the standard GLM-5.2 MAPI.
Workflows are now in Grok Build.
Hand Grok a task a single conversation can't hold: Triage 100+ issues, or review thousands of lines of code in detail.
Build plans the work, runs up to hundreds of agents in parallel, and comes back with one report.
Voice mode now runs on Claude's more capable models and reaches the tools you've connected mid-conversation.
Talk through the hard problems out loud, in many more languages.
Today, @AMD and @cerebras announce a historic partnership.
A new disaggregated inference architecture that combines the best of both worlds:
AMD Helios for world-class prefill performance and the Cerebras Wafer-Scale Engine for the industry’s fastest decode.
For years, AI inference forced a tradeoff between throughput and latency.
Developers don’t want that tradeoff. Neither do users.
The result of our partnership is an entirely new performance envelope for frontier AI: ultra-low latency at massive scale.
Faster AI changes what developers can build.
Richer user experiences.
Faster software development, robotics innovations and scientific discovery.
Entirely new classes of applications that simply weren't possible when inference is slow.
Proud to partner with @LisaSu and the AMD team again.
Today, @AMD and Cerebras introduced a powerful disaggregated inference solution, pairing the right engine to each phase of the inference pipeline.
This is what agentic AI has been waiting for: the fastest production inference at massive scale.
🚨 GPT-5.6 Sol is entering its Cerebras launch window
OpenAI’s fastest frontier model deployment is scheduled to begin this month:
• Up to 750 output tokens per second
• Full GPT-5.6 Sol intelligence
• Initial access limited to selected customers
• Capacity will expand gradually
• Long coding and reasoning outputs could finish in seconds
This could make a frontier model feel almost instant for coding, research and agent tasks.
Are you looking forward to this rolling out and what will you test it on first?
Cursor + Grok 4.5 ran nonstop for the last 24 hours, burned through 214,823,154 tokens, and only moved my usage by 30% (from 35%) on a $20 plan.
Take advantage of the doubled usage limits.
We've just added several new features to Claude Managed Agents.
You can now configure effort levels per agent, seed sessions with events, add up to 500 skills per session, use webhooks for environments + memory stores, and stream events for sub-agents.