Terminal Bench v4 scores
A discussion on reddit.com highlights Terminal Bench v4 scores, suggesting they might better reflect model intelligence than the intelligent index. The rankings appear to align with public perception of various open and closed models. GLM-5.3 leads with 41.9%, followed by GLM-5.3-Flash at 32.8% and DSV4.1-Flash at 26.8%. Other models like Qwen3.8-Flash-Next, DSV4-Pro, Kimi-K3, and DSV4-Flash scored lower, with gemma4-31b at 0.0%.
Why this oneThis discussion uniquely presents Terminal Bench v4 scores, unlike other reports that focus on the intelligent index, offering a different perspective on model intelligence.
Time & source
Times shown in UTC
Display time zone: UTC
Local time zone unavailable; showing UTC.
IngestedOffset at this time: UTC+0Sep 11, 2026, 22:00 UTC
- Ingested
- Sep 11, 2026, 22:00
- Source type
- Dev community
Discussion trend
The percentage is based on collected discussion signal, not new comments or independent people. The curve only compares the same topic across time.
Full text isn't available here.
Read at source →