Anthropic Internally Uses A Model That Is Significantly Better Than Mythos 5, But Has No Plans To Release It
Heat trend
Collecting trend data
The percentage is based on available heat signal, not comment count or independent people.
Claude model activity is surfacing — worth tracking for capability changes, ecosystem impact, and availability.
Anthropic reportedly uses an internal model, "Model 2," that significantly outperforms Mythos 5, scoring 12.5 percentage points higher on CoBench v2.…
https://x.com/kimmonismus/status/2088331147650490748
As has already been expected, Anthropic internally uses a model that is significantly better than Mythos 5, but they have no plans to release it.
https://x.com/daniel_mac8/status/2088344245178716175
RSI is near.
Anthropic tested an unreleased ‘Model 2’ on CoBench v2. CoBench tests a model’s ability to solve historical AI R&D tasks that Anthropic staff solved.
Model 2 scored 12.5 percentage points higher than Mythos 5.
The report estimates a model that scores 85% could replace Anthropic researchers. Only a couple turns of the crank left.
The Singularity is here. 2027 is the takeoff. It’s 4 months away.