Remix.run Logo
▲ prodigycorp 3 hours ago

Incorrect.

Anthropic has admitted to nerfing in the past. There have also been inference bugs. On top of that, model performance changes as they move compute to schwaggier providers as well.

Your chart is wrong.

▲simonw 3 hours ago | parent | next [-]

> Anthropic has admitted to nerfing in the past

Where?

▲QwenGlazer9000 2 hours ago | parent | next [-]

Earlier in march/aprile, there was a regression in Claude code.

Unintentional tbf.

▲prodigycorp 3 hours ago | parent | prev [-]

Man this was last year and some Claude subreddit drama that I can’t furnish offf the top of my head but maybe one of the historians remember it.

▲p-e-w 2 hours ago | parent | next [-]

Noone can seem to remember anything with certainty when asked to actually substantiate these claims.

▲computerex 2 hours ago | parent | next [-]

There was this: https://www.reddit.com/r/Anthropic/comments/1sl5wfh/the_degr...

▲prodigycorp 2 hours ago | parent | prev | next [-]

I’m typing from my phone and im not going to review the semantics of Anthropic’s storied history of performance issues.

It’s not just ant. There are so many small knobs that providers can claim isn’t nerfing but “load management” or “improving user experience”. One example from OpenAI is reducing juice to reduce time to first token.

▲winwang an hour ago | parent [-]

You can just have your agent find the evidence, review it, copypaste it.

▲erinnh 2 hours ago | parent | prev [-]

I mean there is a direct link two comments down from here from 30 minutes before your comment: https://news.ycombinator.com/item?id=49902477

▲p-e-w an hour ago | parent [-]

That’s NOT Anthropic admitting to “nerfing” their model as claimed above (which implies intent), that’s a regression which they quickly fixed.

Christ this forum has become intellectually dishonest.

▲consumer451 3 hours ago | parent | prev [-]

I asked a historian:

Two postmortems, neither quite "admitted to nerfing":

Sept 2025, infra bugs: "A small percentage of Claude Sonnet 4 requests experienced degraded output quality" [0], alongside "We never reduce model quality due to demand, time of day, or server load." [1]

April 2026, Claude Code: default reasoning effort was lowered from high to medium, plus a caching bug and a verbosity prompt. Per Anthropic, "The models themselves didn't regress, and the Claude API was not affected." [2]

So users were right that quality dropped, but the confirmed causes were bugs and a product default, not deliberate model degradation.

[0] https://status.claude.com/incidents/72f99lh1cj2c

[1] https://anthropic.com/engineering/a-postmortem-of-three-rece...

[2] https://texxr.com/handle/claudedevs

source: https://claude.ai/share/4435bbcf-d6df-44a0-b1db-f08a11858bc2

▲what an hour ago | parent [-]

> bugs

There are no bugs, just happy little accidents.

▲consumer451 an hour ago | parent [-]

u/bcherny does sound a bit like Bob Ross now that you mention it.

▲johnfn 3 hours ago | parent | prev | next [-]

Sorry, you are correct - I modified my original post. I get frustrated every time there's a model release and 1 week later everyone is saying NERF! NERF! 99.9% of the time these people are wrong, but you are right that it's technically not 100% due to a few edge cases.

I am more skeptical about the compute provider claim - do you have any evidence of that?

▲r_lee 2 hours ago | parent [-]

I noticed that a few weeks back 5.6 sol would regularly glitch out and start speeding random words or loop and then the next day it'd be fine

and there's sometimes just huge floods of complaints from people all of a sudden, which is pretty unlikely to be a coincidence

▲swader999 3 hours ago | parent | prev [-]

Right, and it would be simple to un-nerf or shadow nerf by any kind of angle they want.