Xai Rate Limits
xAI enforces per-team rate limits on the synchronous API (requests per minute and tokens per minute) that vary by model and account tier. The Batch API does not count toward synchronous rate limits. Specific per-model limits are not reconciled in this artifact; consult the xAI Console for the active limits on your team.
Xai Rate Limits is the machine-readable rate-limit profile for xAI on the APIs.io network, conforming to the API Commons Rate Limits specification.
It captures 3 rate-limit definitions, measuring requests and tokens.
The profile also includes 2 backoff/retry policies defined and response codes documented for throttled.
Tagged areas include AI, LLM, Foundation Models, Grok, and Generative AI.
Limits
Policies
Sources
Work with this as data
Every rate limit here is available over the APIs.io API and to AI agents over MCP.