Anthropic Internally Uses A Model That Is Significantly Better Than Mythos 5, But Has No Plans To Release It
热度趋势
趋势数据积累中
百分比基于当前可用热度信号,而非评论数或独立用户人数。
Claude 相关模型动态已经出现,适合跟踪能力变化、生态影响和后续可用性。
据报道,Anthropic内部使用一个名为“Model 2”的模型,其性能显著优于Mythos 5,在CoBench v2测试中得分高出12.5个百分点。尽管该模型能力先进,一份报告估计得分达到85%的模型可能取代Anthropic的研究人员,但目前Anthropic没有发布该模型的计划。…
https://x.com/kimmonismus/status/2088331147650490748
As has already been expected, Anthropic internally uses a model that is significantly better than Mythos 5, but they have no plans to release it.
https://x.com/daniel_mac8/status/2088344245178716175
RSI is near.
Anthropic tested an unreleased ‘Model 2’ on CoBench v2. CoBench tests a model’s ability to solve historical AI R&D tasks that Anthropic staff solved.
Model 2 scored 12.5 percentage points higher than Mythos 5.
The report estimates a model that scores 85% could replace Anthropic researchers. Only a couple turns of the crank left.
The Singularity is here. 2027 is the takeoff. It’s 4 months away.