Segmind · Rate Limits

Segmind Rate Limits

Published rate limits for the Segmind AI Gateway, harvested from the provider's own rate-limits documentation, pricing page and authentication guide on 2026-08-27. This replaces the 2026-05-04 scaffold, whose tier numbers were placeholders.

Segmind Rate Limits is the machine-readable rate-limit profile for Segmind on the APIs.io network, conforming to the API Commons Rate Limits specification.

It captures 7 rate-limit definitions, across the free, professional, business, scale, and enterprise tiers, measuring requests_per_minute.

The profile also includes 3 backoff/retry policies defined and response codes documented for throttled, quotaExceeded, and note.

Tagged areas include Rate Limiting, Quotas, and Throttling.

7 Limits Throttle: 429 Quota: 406
Rate LimitingQuotasThrottling

Limits

Docs rate-limits page — Free api-key
requests_per_minute · minute
5
Docs rate-limits page — Pro api-key
requests_per_minute · minute
50
Pricing page — Flexible (pay as you go) account
requests_per_minute · minute
60
Pricing page — Pro account
requests_per_minute · minute
120
Pricing page — Business account
requests_per_minute · minute
500
Pricing page — Scale (pooled) account
requests_per_minute · minute
1000
Enterprise — custom contract
requests_per_minute · minute

Policies

Retry guidance
The authentication docs tell clients to implement retry with exponential backoff and to handle token expiration gracefully. No Retry-After header is documented to back that up, so the backoff is entirely client-side.
Dedicated endpoints bypass the shared limit
Dedicated GPU endpoints are billed per GPU-hour or per GPU-second rather than per call, and the pricing page describes them as "Unlimited API requests with dedicated endpoints". They are a separate control plane on api.spotprod.segmind.com.
Why limits exist
Segmind states the purpose as security (abuse mitigation), fair access (no single consumer monopolising capacity), and performance optimisation under spikes.

Work with this as data

Every rate limit here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for rate limits

4 MCP tools reach this
  • find_rate_limitsBrowse and filter every rate limit in the catalog.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This rate limit
curl "https://apis.io/api/v1/rate-limits/segmind-rate-limits"
All rate limits
curl "https://apis.io/api/v1/rate-limits?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no form to fill in. Signing in shares your email address with us — we store it to create your key and to recognise you if you sign in with another provider. See our Privacy Policy and Terms.

A second provider on the same verified email joins the account you already have.