
Cognition's SWE-2 Gets Within a Point of the Frontier for 64% Less
Cognition says SWE-2 scores 50.0% on FrontierCode 1.1 Main, a point behind Fable 5.1, at 64% lower cost — built on an open Chinese base model.
Seung Jung·
5m
Cognition says SWE-2 scores 50.0% on FrontierCode 1.1 Main, a point behind Fable 5.1, at 64% lower cost — built on an open Chinese base model.

IBM's Granite 4.2 ships 3B, 8B and 30B dense reasoning models under Apache 2.0 with a 512K context window and an agentic RL stage for the larger two.

The IOL-AI Challenge had the official Linguistics Olympiad jury grade machine entries. Claude Opus 4.8 hit gold-medal marks; scale did not predict results.