Introducing SWE-2, our closest model yet to the frontier. On...

In a single RL run, we are able to push the Pareto curve while preserving its shape: medium effort becomes both cheaper & smarter; max effort learns to use more tokens & turns to achieve the highest scores.
Our team at Cognition uses it for feature development, debugging, and even complex visualizations of novel math:

(singularity at t = 1 with viscosity matching water, unit distance ~ 1 mm. leading and first order background terms)
x.com/OpenAI/status/…
Read more about how we trained SWE-2:
cognition.com/blog/swe-2


