Inflection Rate Limits
Inflection AI's Developer API exposes a usage dashboard at developers.inflection.ai/usage that surfaces account-level rate-limit consumption. The exact RPM / TPM values per model and tier are not publicly documented and are pending reconciliation; in practice they are set per contract.
Inflection Rate Limits is the machine-readable rate-limit profile for Inflection AI on the APIs.io network, conforming to the API Commons Rate Limits specification.
It captures 3 rate-limit definitions, measuring requests and tokens.
The profile also includes 2 backoff/retry policies defined and response codes documented for throttled.
Tagged areas include AI, LLM, Personal AI, Pi, and Foundation Models.
Limits
Policies
Sources
Work with this as data
Every rate limit here is available over the APIs.io API and to AI agents over MCP.