HNHacker News·
暂不在当前实时榜单
Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra
Cognition has launched its new SWE-2 model, which demonstrates strong performance in coding benchmarks, rivaling models like Fable 5.1 and GPT-Astra. The SWE-2 model achieved 50.0% on Main 50.0 %, 73.0% on DeepSWE 1.1, 92.8% on Terminal-Bench 2.1, and 27.3% on Terminal-Bench 4. Cognition also shared that it uses a length-weighted reward baseline, introduced since SWE-1.6, to stabilize training and reduce gradient variance.
This report is the first to detail Cognition's SWE-2 model, which outperforms SWE-1.7 across all listed benchmarks and introduces a new reward baseline.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年9月10日 17:00 UTC
- 收录
- 2026年9月10日 17:00
- 来源类型
- 未分类
- 判定依据
- 热度约为该来源近期上榜条目中位水平的 6.6 倍
- 指标对比
- 309 vs 中位 47(20 条基线样本)
- 检出时间
- 09/10 21:00
本站未收录正文。
前往源站阅读 →