跳到正文
RCreddit.com·
暂不在当前实时榜单

Terminal Bench v4 scores

AI 摘要

A discussion on reddit.com highlights Terminal Bench v4 scores, suggesting they might better reflect model intelligence than the intelligent index. The rankings appear to align with public perception of various open and closed models. GLM-5.3 leads with 41.9%, followed by GLM-5.3-Flash at 32.8% and DSV4.1-Flash at 26.8%. Other models like Qwen3.8-Flash-Next, DSV4-Pro, Kimi-K3, and DSV4-Flash scored lower, with gemma4-31b at 0.0%.

为什么是这条

This discussion uniquely presents Terminal Bench v4 scores, unlike other reports that focus on the intelligent index, offering a different perspective on model intelligence.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月11日 22:00 UTC

收录
2026年9月11日 22:00
来源类型
开发者社区

讨论趋势

暂无对比
最近 24 小时与此前 24 小时的快照均值对比 · 7 天曲线

百分比基于采集到的讨论信号,不代表新增评论数或独立参与人数。曲线仅用于同一话题在不同时段的比较。

正文

本站未收录正文。

前往源站阅读 →
来源·reddit.com