Open Webui Rate Limits
Open WebUI does not impose project-level API rate limits. Effective limits are determined by (1) the upstream LLM backend's limits (Ollama concurrency, or OpenAI/Anthropic RPM caps) and (2) any reverse-proxy or admin throttling configured in the deployment. Standard HTTP semantics apply.
Open Webui Rate Limits is the machine-readable rate-limit profile for Open WebUI on the APIs.io network, conforming to the API Commons Rate Limits specification.
It captures 2 rate-limit definitions, measuring n/a and requests.
The profile also includes 2 backoff/retry policies defined and response codes documented for throttled.
Tagged areas include LLM, Open Source, Self-Hosted, Ollama, and Chat UI.
Limits
Policies
Sources
Work with this as data
Every rate limit here is available over the APIs.io API and to AI agents over MCP.