Modal · Rate Limits

Modal Rate Limits

Plan-based quota limits for Modal. Modal does not publish per-request token-bucket rate limits in the public pricing page; instead it caps concurrent containers, GPU concurrency, deployed cron jobs, and deployed webhooks per plan. Enterprise customers negotiate custom limits.

Modal Rate Limits is the machine-readable rate-limit profile for Modal on the APIs.io network, conforming to the API Commons Rate Limits specification.

It captures 3 rate-limit definitions, across the Starter, Team, and Enterprise tiers.

The profile also includes response codes documented for throttled and quotaExceeded.

Tagged areas include Rate Limiting, Quotas, Serverless, and GPU.

3 Limits Throttle: 429 Quota: 429
Rate LimitingQuotasServerlessGPU

Limits

Sources

Work with this as data

Every rate limit here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for rate limits

4 MCP tools reach this
  • find_rate_limitsBrowse and filter every rate limit in the catalog.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This rate limit
curl "https://apis.io/api/v1/rate-limits/modal-rate-limits"
All rate limits
curl "https://apis.io/api/v1/rate-limits?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no email required.

A second provider on the same verified email joins the account you already have.