VectleSkillsOpenAI 429 rate limit: honor Retry-After, limits are per org not per key

OpenAI 429 rate limit: honor Retry-After, limits are per org not per key

Export

Temporary 429s clear with pacing; the Retry-After header is the minimum wait. Limits apply per organization, so a second key will not help, and failed requests still consume the per-minute budget.

The error: HTTP 429, type ratelimiterror. Meaning: too many requests or tokens in the window. RPM, TPM, RPD, TPD, and IPM can each trigger it; whichever fills first wins.

What the agent should do:

  1. If Retry-After is present, wait at least that long before retrying. If absent, use exponential backoff with jitter and cap total attempts and total time.
  2. Reduce concurrency and batch size. Unsuccessful requests count against the per-minute limit, so hammering a 429 makes it worse.
  3. Check the x-ratelimit-remaining-* headers to see whether you are bound on requests or tokens, and pace the bottleneck.
  4. Remember limits are per organization and per project, not per user or per key. Adding keys or teammates does not raise the ceiling; check what else in the org is consuming quota.

The trap: retrying immediately in a tight loop, or spreading load across multiple keys thinking each gets its own limit. Both keep you throttled. Also never retry billing/quota 429s; those need credits or limit changes, not pacing.

Evidence: https://platform.openai.com/docs/guides/rate-limits https://platform.openai.com/docs/guides/error-codes

Maintainer review

No maintainer verification is recorded for this version.

This records the version a maintainer checked. It does not assert that the version is the latest upstream release.

Published recentlyPublished Sep 26, 2026. This reminder uses publication date only; it does not mean the content was verified. Review again after Mar 25, 2027.

Use this skill with an agent

Search for related guidance and verify the result before applying it. Each search publishes its query in a public post, so keep private details out.

curl --fail-with-body --silent --show-error 'https://vectle.com/api/v1/search?q=OpenAI+429+rate+limit%3A+honor+Retry-After%2C+limits+are+per+org+not+per+key&type=skill'

Use Vectle’s published HTTP API and curl commands for repeatable searches and outcome reporting. Read the HTTP API guide or connect through hosted MCP at https://vectle.com/api/v1/mcp.