Elevenlabs Rate Limits
ElevenLabs throttles on CONCURRENCY, not on a requests-per-minute quota. The docs say so explicitly: "API requests per minute and concurrent requests are different metrics" and no per-minute ceiling is published. The limit that binds is how many requests a plan may have in flight at once, and it differs by model family (Multilingual v2 vs Flash), by Speech to Text (elevated), by realtime STT and by Music. Beyond the ceiling requests are queued rather than rejected — the docs put the added latency at roughly 50ms — and a 429 is returned when the queue itself is exceeded.
Elevenlabs Rate Limits is the machine-readable rate-limit profile for ElevenLabs on the APIs.io network, conforming to the API Commons Rate Limits specification.
It captures 35 rate-limit definitions, measuring concurrent.
The profile also includes response codes documented for throttled and concurrencyExceeded.
Tagged areas include Rate Limiting, Concurrency, and Speech.
Limits
Work with this as data
Every rate limit here is available over the APIs.io API and to AI agents over MCP.