Amazon Nova Rate Limits
Published Amazon Bedrock service quotas that govern Amazon Nova model invocation, harvested from the AWS General Reference quota tables on 2026-09-01. This file REPLACES an API Evangelist scaffold dated 2026-05-04 that asserted invented per-key request quotas and X-RateLimit-* headers; none of those existed. Amazon Nova is not rate limited per API key at all — it is quota'd per AWS account, per region, per MODEL, in requests-per-minute and tokens-per-minute.
Amazon Nova Rate Limits is the machine-readable rate-limit profile for Amazon Nova on the APIs.io network, conforming to the API Commons Rate Limits specification.
It captures 20 rate-limit definitions, measuring requests_per_minute, tokens_per_minute, and concurrent_requests.
The profile also includes 3 backoff/retry policies defined and response codes documented for throttled, quotaExceeded, and serviceUnavailable.
Tagged areas include Foundation Models, Rate Limiting, Quotas, and Throttling.
Limits
Policies
Work with this as data
Every rate limit here is available over the APIs.io API and to AI agents over MCP.