Pinecone · Rate Limits
Pinecone Rate Limits
Pinecone Serverless concurrency-based rate limits per index.
Pinecone Rate Limits is the machine-readable rate-limit profile for Pinecone on the APIs.io network, conforming to the API Commons Rate Limits specification.
It captures 4 rate-limit definitions, measuring queries_per_second, concurrent, concurrent_imports, and vectors.
The profile also includes response codes documented for throttled.
Tagged areas include Rate Limiting and Vector Database.
4 Limits
Throttle: 429
Rate LimitingVector Database
Limits
Index throughput index
scales with read units
Concurrent operations index
varies
Bulk import index
10
Records per upsert request
1000
Sources
Work with this as data
Every rate limit here is available over the APIs.io API and to AI agents over MCP.