xAI Rate Limits by Tier

What are xAI's API rate limits?

xAI's entry tier (Default) allows an unpublished number of requests/min and an unpublished number of tokens/min for documented default/model family. Limits scale up through Default as cumulative spend and account age increase — see the full table below, verified 2026-08-15.

Verified 2026-08-15 source

Limits by tier

TierQualificationModel classRPMTPMRPDConcurrent
DefaultAccount/project limit shown in the xAI consoledocumented default/model family

— means not documented by xAI, never a guess.

What this means for your workload

Classification at volume: 116 calls/min and 60,320 tokens/min at the production profile.

This provider publishes no numeric cap for this workload; check the account console before launch.

Response headers

retry-afterSeconds to wait before retrying, when supplied with a 429
rate-limit response headersProvider-specific remaining and reset counters when documented

When you exceed the limit

xAI returns HTTP 429.

Use the limit shown for the selected model and project, then back off.

FAQ

What happens when I exceed xAI's rate limit?

xAI returns HTTP 429. Use the limit shown for the selected model and project, then back off.

How do I request a rate limit increase on xAI?

Request one from the account dashboard: https://console.x.ai

xAI provider hubGet a xAI API key