Thundercompute Rate Limits
Thunder Compute does not publish explicit numeric request-per-minute rate limits for its REST API. The primary constraints are account-level resource quotas - the number and type of GPU instances an account may provision concurrently - which are governed by available GPU capacity, the account's tier/standing, and payment status rather than by API call throttling. New accounts may have lower concurrent-instance and GPU-type quotas that are raised on request. Specific numeric limits are not reconciled in this artifact.
Thundercompute Rate Limits is the machine-readable rate-limit profile for Thunder Compute on the APIs.io network, conforming to the API Commons Rate Limits specification.
It captures 3 rate-limit definitions, measuring instances, gpus, and requests.
The profile also includes 2 backoff/retry policies defined and response codes documented for throttled.
Tagged areas include GPU, Cloud, Infrastructure, AI, and Compute.
Limits
Policies
Sources
Work with this as data
Every rate limit here is available over the APIs.io API and to AI agents over MCP.