API rate limits
API endpoints (/api/v2/generate and /v1/chat/completions) use a single limit: 30 requests per minute per account. The limit is the same for all plans.
Bot rate limits by plan
This table is for the Telegram bot, not for API responses:
Bot limits are applied per user account, not per API key.
Rate Limit Headers
OpenAI-compatible API responses (/v1/chat/completions) include only:
X-RateLimit-Remaining and X-RateLimit-Reset are not returned in API responses.
Handling rate limit errors
429 is returned.
Best Practices
Implement exponential backoff
Implement exponential backoff
Don’t hammer the API after a 429. Wait, then retry with increasing delays:
Monitor limits proactively
Monitor limits proactively
Use the API docs values (
30 requests/minute) to size your retry schedule; these are not per-plan for API.Batch where possible
Batch where possible
For bulk tasks (e.g. generating 100 images), spread requests over time.

