Rreddit·
Archived topic · 归档话题,来源已停止追踪
DeepSWE just added the gpt-5.6 models to their benchmark. I hope you guys don't get too used to Claude Code as your only coding agent. Chart is marked NSFW due to the grotesque violence.
DeepSWE's latest benchmark now includes gpt-5.6 models, potentially challenging Claude Code's dominance as a coding agent. The chart, marked NSFW, shows various models' performance and average cost per task. Gpt-5.6-sol is highlighted as the most efficient, while gpt-5.4 XHIGH and claude-fable-5 HIGH are among the higher-cost options. This update suggests a shift in the competitive landscape for coding agents.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年7月10日 04:00 UTC
- 收录
- 2026年7月10日 04:00
- 来源类型
- 未分类