In the News: September 10, 2026 (Extra 4)
On Cognition's benchmark, SWE-2 scores within one point of Anthropic's Fable 5.1 at 64 percent lower cost and reaches its first code edit in fewer than half the steps of SWE-1.7.
Extra edition
Machine-readable
Download Markdown
Story
SWE-2 scores within one point of Fable 5.1 on Cognition's benchmark
Cognition released SWE-2, a coding model post-trained from Moonshot's 2.8-trillion-parameter Kimi K3. On the company's FrontierCode 1.1 Main benchmark, SWE-2 scores 50.0 percent.
Read story →