After Astra's stealth nerf last night, we really need benchmarks to do a re-bench 1 week after any model release. This is ridiculous.
A user on reddit.com expressed frustration over a "stealth nerf" to the Astra model, claiming its capabilities significantly diminished overnight. They noted that Astra, which previously excelled at "1 shotting AAA game level models," now produces "completely terrible" results even after "reprompting on xhigh effort several times." The user criticized "frontier labs" for quietly reducing model capabilities to save compute, especially after many users purchased "Pro subscriptions" based on initial performance, suggesting the "FTC should get involved."
Why this oneThis report highlights a user's direct experience of a model's performance degradation, unlike many discussions that focus on theoretical capabilities or benchmark scores.
Time & source
Times shown in UTC
Display time zone: UTC
Local time zone unavailable; showing UTC.
IngestedOffset at this time: UTC+0Sep 11, 2026, 04:00 UTC
- Ingested
- Sep 11, 2026, 04:00
- Source type
- Dev community
Discussion trend
The percentage is based on collected discussion signal, not new comments or independent people. The curve only compares the same topic across time.
Full text isn't available here.
Read at source →