AWS activates Project Rainier: One of the world’s largest AI compute clusters comes online.
~500,000 Trainium2 chips, and Anthropic is scaling Claude to >1,000,000 chips by Dec-25, a huge jump in training and inference capacity.
AWS connected multiple US data centers into one UltraCluster so Anthropic can train larger Claude models and handle longer context and heavier workloads without slowing down.
Each Trn2 UltraServer links 64 Trainium2 chips through NeuronLink inside the node, and EFA networking connects those nodes across buildings, cutting latency and keeping the cluster flexible for massive scaling.
Trainium2 is optimized for matrix and tensor math with HBM3 memory, giving it extremely high bandwidth so huge batches and long sequences can be processed without waiting for data transfer.
The UltraServers act as powerful single compute units inside racks, while the UltraCluster spreads training across tens of thousands of these servers, using parallel processing to handle giant models efficiently.
AWS says Project Rainier is its largest training platform ever, delivering >5x compute than what Anthropic used before, allowing faster model training and easier large-scale experiments.
For energy use, AWS reports a 0.15 L/kWh water usage efficiency, matching 100% renewable power and adding nuclear and battery investments to keep growing while staying within its 2040 net-zero goal.
---
aboutamazon. com/news/aws/aws-project-rainier-ai-trainium-chips-compute-cluster
Snyk's #CTF, #FetchTheFlag is back — with an exciting guest! Drumroll, please... 🥁
@_johnhammond and Snyk are teaming up! Get ready to tackle 30 challenges and compete to win a Nintendo Switch on October 27th. 👾 Register today: https://t.co/64XQ4MNYOz
https://t.co/PIqn1yzvvg
@O2 your service has been down in my area for almost a week now. Your website confirms there’s an issue but I would expect this to be fixed quickly, have DM’d you my postcode. Will customers get a credit on their accounts for this service disruption?