Introducing SWE-2, our closest model yet to the frontier.
On leading evals, it scores on par with recent frontier models – at up to 70% lower cost.
We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & cost.
@cohere my understanding is that megakernel hits diminishing returns as the batch size increases, can you post the results of > 16, or > 32 numbers?