跳到正文
HNHacker News·
暂不在当前实时榜单

Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra

AI 摘要

Cognition has launched its new SWE-2 model, which demonstrates strong performance in coding benchmarks, rivaling models like Fable 5.1 and GPT-Astra. The SWE-2 model achieved 50.0% on Main 50.0 %, 73.0% on DeepSWE 1.1, 92.8% on Terminal-Bench 2.1, and 27.3% on Terminal-Bench 4. Cognition also shared that it uses a length-weighted reward baseline, introduced since SWE-1.6, to stabilize training and reduce gradient variance.

为什么是这条

This report is the first to detail Cognition's SWE-2 model, which outperforms SWE-1.7 across all listed benchmarks and introduces a new reward baseline.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月10日 17:00 UTC

收录
2026年9月10日 17:00
来源类型
未分类
爆款判定
判定依据
热度约为该来源近期上榜条目中位水平的 6.6 倍
指标对比
309 vs 中位 47(20 条基线样本)
检出时间
09/10 21:00

本站未收录正文。

前往源站阅读 →
来源·Hacker News·cognition.com