·
Archived topic · 归档话题,来源已停止追踪
Kimi-K3 on HuggingFace
Kimi-K3 on HuggingFace presents various benchmark results, including Coding and Agentic capabilities. For Coding, benchmarks like DeepSWE, ProgramBench, Terminal-Bench 2.1, FrontierSWE, SWE-Marathon, PostTrainBench, MLS-Bench-Lite, SciCode, and Kimi Code Bench 2.0 are listed with their respective scores. Agentic capabilities are evaluated using BrowseComp. Additionally, PerceptionBench is mentioned as an in-house benchmark focusing on atomic visual perception capabilities.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年7月27日 11:00 UTC
- 收录
- 2026年7月27日 11:00
- 来源类型
- 未分类