Sarvam AI · Rate Limits

Sarvam Ai Rate Limits

Sarvam AI enforces per-account rate limits expressed primarily as requests per minute (RPM), applied per product (chat LLM, speech-to-text, text-to-speech, translate). Limits are enforced at the account level - all API keys on an account share the same pool - and rise with plan tier (Starter, Pro, Business, Enterprise). The Starter plan allows 60 RPM for the chat LLM API. Specific per-product values for higher tiers are not reconciled in this artifact.

Sarvam Ai Rate Limits is the machine-readable rate-limit profile for Sarvam AI on the APIs.io network, conforming to the API Commons Rate Limits specification.

It captures 4 rate-limit definitions, measuring requests.

The profile also includes 3 backoff/retry policies defined and response codes documented for throttled.

Tagged areas include AI, LLM, Speech to Text, Text to Speech, and Translation.

4 Limits Throttle: 429
AILLMSpeech to TextText to SpeechTranslationIndian LanguagesRate LimitingQuotasThrottling

Limits

Requests Per Minute (RPM) - Chat LLM account
requests
60 on Starter; higher on paid tiers
Per-account RPM for the chat completions API; shared across all keys.
Requests Per Minute (RPM) - Speech-to-Text account
requests
see provider dashboard
Per-account RPM for speech-to-text and speech-to-text-translate; varies by tier.
Requests Per Minute (RPM) - Text-to-Speech account
requests
see provider dashboard
Per-account RPM for text-to-speech; varies by tier.
Requests Per Minute (RPM) - Translate / Transliterate account
requests
see provider dashboard
Per-account RPM for translate, transliterate, and language ID; varies by tier.

Policies

Account-Level Enforcement
Limits apply to the account as a whole; all API keys share the same rate-limit pool.
Tiered Limits
Limits raise as accounts move from Starter to Pro, Business, and Enterprise plans.
Backoff Strategy
Clients should implement exponential backoff with jitter and honor Retry-After on 429 responses.

Sources

Work with this as data

Every rate limit here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for rate limits

4 MCP tools reach this
  • find_rate_limitsBrowse and filter every rate limit in the catalog.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This rate limit
curl "https://apis.io/api/v1/rate-limits/sarvam-ai-rate-limits"
All rate limits
curl "https://apis.io/api/v1/rate-limits?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no form to fill in. Signing in shares your email address with us — we store it to create your key and to recognise you if you sign in with another provider. See our Privacy Policy and Terms.

A second provider on the same verified email joins the account you already have.