HNHacker News·
Not on the current live radar
Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra
Cognition has launched its new SWE-2 model, which demonstrates strong performance in coding benchmarks, rivaling models like Fable 5.1 and GPT-Astra. The SWE-2 model achieved 50.0% on Main 50.0 %, 73.0% on DeepSWE 1.1, 92.8% on Terminal-Bench 2.1, and 27.3% on Terminal-Bench 4. Cognition also shared that it uses a length-weighted reward baseline, introduced since SWE-1.6, to stabilize training and reduce gradient variance.
This report is the first to detail Cognition's SWE-2 model, which outperforms SWE-1.7 across all listed benchmarks and introduces a new reward baseline.
Time & source
Times shown in UTC
Display time zone: UTC
Local time zone unavailable; showing UTC.
IngestedOffset at this time: UTC+0Sep 10, 2026, 17:00 UTC
- Ingested
- Sep 10, 2026, 17:00
- Source type
- Unclassified
- Basis
- Running about 6.6× the median of this source's recent listed items
- Metric comparison
- 309 vs median 47 (20 baseline samples)
- Detected
- 09/10, 21:00
Full text isn't available here.
Read at source →