To show what these optimizations make possible, we’re partnering with Adaptyv Bio on a protein design competition. Together, we’ll be experimentally validating over 5,000 designs.
We're providing up to $1 million in Claude credits plus funding alongside Adaptyv for experimental
Runtime speaker lineup is live!
We're bringing together experts covering AI infrastructure, applications of AI in science and robotics, the future of software engineering, and more.
Frontier models are simply too expensive and slow for the majority of use cases, so we see models like SWE-2 becoming the daily driver for most.
Training trillion-parameter coding agents at scale isn't easy though: typically, each step launches thousands of rollouts, each with
Introducing SWE-2, our closest model yet to the frontier.
On leading evals, it scores on par with recent frontier models – at up to 70% lower cost.
We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & cost.