Remix.run Logo
tvbusy 2 hours ago

Struggling with demands is a performance problem, not a reliability problem.

swe_dima 2 hours ago | parent | next [-]

overly high demand can often put a system in an unstable state

therein 2 hours ago | parent [-]

Only if you're incapable of throttling to assure quality of service. Are you saying they have so much demand that even their load balancers were overloaded?

slopinthebag an hour ago | parent [-]

There is no point having this discussion. Some people have had their brains genuinely broken by Anthropic and will defend them to the last breath, despite having no clue what they are talking about.

36799753298 22 minutes ago | parent [-]

I'll leave it to the retards to confidently claim load has no bearing on the stability of a distributed system.

chrisjj 5 minutes ago | parent [-]

[delayed]

ACCount37 an hour ago | parent | prev [-]

It is a reliability problem - because if the API ends up rejecting 30% of the incoming queries, no one cares if it's 500 Internal Server Error or 529 Overloaded.

Having your infrastructure at the knife's edge of load to capacity also means you have no redundancy when something fails.

cherioo an hour ago | parent [-]

If this incident had happened during the day I could buy that argument. But it is happening in the dead of the night. Are there really that much people scheduling claude runs over-night, taking more capacity than day-time? (assume most claude users are west coast programmers)

My other guess is they are scheduling training runs and their capacity isolation is bad.

cbg0 29 minutes ago | parent [-]

The world is more than just the Americas.