
OpenAI Splits the Rate-Limit 429: What 'slow_down' vs 'server_is_overloaded' Mean for Your Retry Policy
OpenAI now returns distinct error codes for two previously identical 429s — traffic that ramps too fast gets slow_down, model saturation gets a 503. Your retry logic must treat them differently, and most AI gateway clients currently don't.


