Portkey Rate Limits
Portkey is an LLM gateway whose runtime quotas are dominated by the upstream provider being proxied (OpenAI, Anthropic, Bedrock, etc.); Portkey itself exposes plan-bound caps on recorded logs (10k/month on Developer, 100k/month on Production with overage to 3M, 10M+ on Enterprise) rather than per-second request throttling. Enterprise customers can configure granular budget and rate limits per virtual key and workspace. Concrete numeric request-per-second ceilings are not published on the public docs site at the time of writing.
Portkey Rate Limits is the machine-readable rate-limit profile for Portkey on the APIs.io network, conforming to the API Commons Rate Limits specification.
It captures 5 rate-limit definitions, measuring requests_per_month and varies.
The profile also includes 4 backoff/retry policies defined and response codes documented for throttled, quotaExceeded, and serviceUnavailable.
Tagged areas include AI Gateways, Governance, Observability, and Rate Limiting.
Limits
Policies
Sources
- https://portkey.ai/pricing
- https://portkey.ai/docs/api-reference/inference-api/introduction
- https://portkey.ai/docs/product/enterprise-offering/private-cloud-deployments
Work with this as data
Every rate limit here is available over the APIs.io API and to AI agents over MCP.
MCP server
One button, every client — Claude, Cursor, VS Code and the rest.
https://apis.io/mcp
Tools for rate limits
4 MCP tools reach this
find_rate_limitsBrowse and filter every rate limit in the catalog.apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.resolveTurn a domain, URL or GitHub org into the provider it belongs to.find_cohortsEvery scored population of providers in the catalog.
Call it yourself
curl for this page
curl "https://apis.io/api/v1/rate-limits/portkey-rate-limits"
curl "https://apis.io/api/v1/rate-limits?limit=25"
Discovery needs no key. Ratings and market analysis are Pro.