返回
RCreddit.com
21
·9小时前·开发者社区 · RSS

Anthropic Internally Uses A Model That Is Significantly Better Than Mythos 5, But Has No Plans To Release It

查看原文
Claude模型发布

热度趋势

趋势数据积累中

百分比基于当前可用热度信号,而非评论数或独立用户人数。

推荐理由

Claude 相关模型动态已经出现,适合跟踪能力变化、生态影响和后续可用性。

AI 摘要

据报道,Anthropic内部使用一个名为“Model 2”的模型,其性能显著优于Mythos 5,在CoBench v2测试中得分高出12.5个百分点。尽管该模型能力先进,一份报告估计得分达到85%的模型可能取代Anthropic的研究人员,但目前Anthropic没有发布该模型的计划。…

https://x.com/kimmonismus/status/2088331147650490748

As has already been expected, Anthropic internally uses a model that is significantly better than Mythos 5, but they have no plans to release it.

https://x.com/daniel_mac8/status/2088344245178716175

RSI is near.

Anthropic tested an unreleased ‘Model 2’ on CoBench v2. CoBench tests a model’s ability to solve historical AI R&D tasks that Anthropic staff solved.

Model 2 scored 12.5 percentage points higher than Mythos 5.

The report estimates a model that scores 85% could replace Anthropic researchers. Only a couple turns of the crank left.

The Singularity is here. 2027 is the takeoff. It’s 4 months away.

Anthropic Internally Uses A Model That Is Significantly Better Than Mythos 5, But Has No Plans To Release It · BuzzRadr