Skip to main content
LLM Gateway is rate limited per model, measured as requests within a 60-second window.
Need a higher rate limit?If you need a higher rate limit, contact our support team.

Response headers

If you exceed the limit, the API responds with a 429 status code. To see your remaining quota, check the following response headers: If the response doesn’t include X-RateLimit headers, the endpoint doesn’t have rate limits.