Skip to main content

API rate limits

API endpoints (/api/v2/generate and /v1/chat/completions) use a single limit: 30 requests per minute per account. The limit is the same for all plans.

Bot rate limits by plan

This table is for the Telegram bot, not for API responses: Bot limits are applied per user account, not per API key.

Rate Limit Headers

OpenAI-compatible API responses (/v1/chat/completions) include only:
X-RateLimit-Remaining and X-RateLimit-Reset are not returned in API responses.

Handling rate limit errors

When the limit is exceeded, status 429 is returned.

Best Practices

Don’t hammer the API after a 429. Wait, then retry with increasing delays:
Use the API docs values (30 requests/minute) to size your retry schedule; these are not per-plan for API.
For bulk tasks (e.g. generating 100 images), spread requests over time.

Auth Endpoint Limits

Authentication endpoints have stricter rate limits to prevent abuse: