π¨ BREAKING: Super excited to announce we're deploying the world's fastest inference with @Cerebras.
Talk to any developer and they're excited to build with 20x faster AI... the problem is there's almost no compute available.
We're here to solve that.
Using GPUs for prefill - it's now more affordable than ever too.
Thank you to the whole Cerebras team and excited to grow this partnership.
First tokens live Q127 π
@sonnet_xu We are firm believers that every single one of these matches / chips will be sold out for years. There is not enough decode silicon in production to keep up with demand for fast inference
π¨ BREAKING: Super excited to announce we're deploying the world's fastest inference with @Cerebras.
Talk to any developer and they're excited to build with 20x faster AI... the problem is there's almost no compute available.
We're here to solve that.
Using GPUs for prefill - it's now more affordable than ever too.
Thank you to the whole Cerebras team and excited to grow this partnership.
First tokens live Q127 π