跳到正文
RCreddit.com·

Why are the SOTA open-weight models scoring (relatively) low scores on AA-Omniscience Index

AI 摘要

Reddit 社区正在讨论为什么最先进(SOTA)的开源模型在 AA-Omniscience 指数上得分相对较低。用户对这些分数感到惊讶,指出它们明显低于 Gemini-3.* flash 模型的得分。这一观察引发了关于人工智能开发社区内性能指标和比较评估的讨论。

时间与来源
发布
09/07 14:09 UTC+0
收录
09/07 21:00 UTC+0
来源类型
开发者社区
档位
社区
信源状态
正常

档位是按信源手工设定的编辑判断,不是逐条打分。

正文 · RSS 全文

https://preview.redd.it/qwmh23nfr3oh1.png?width=1062&format=png&auto=webp&s=4741852dc3f5e7d86ae85281b089b2a325a1804a

I mean they aren't that low but seeing them much lower than Gemini-3.* flash surprises me

来源·reddit.com·RSS 全文