Lamini · Rate Limits

Lamini Rate Limits

The Lamini Platform meters inference and tuning usage per account and is governed primarily by available credit / spend on the On-Demand tier and by reserved GPU capacity on Enterprise. Concurrent inference requests and tuning jobs are bounded by account capacity rather than fixed published per-minute request quotas. Specific numeric limits are not publicly documented and are not reconciled in this artifact.

Lamini Rate Limits is the machine-readable rate-limit profile for Lamini on the APIs.io network, conforming to the API Commons Rate Limits specification.

It captures 4 rate-limit definitions, measuring requests, jobs, steps, and usd.

The profile also includes 2 backoff/retry policies defined and response codes documented for throttled.

Tagged areas include AI, LLM, Fine-Tuning, Memory Tuning, and Inference.

4 Limits Throttle: 429
AILLMFine-TuningMemory TuningInferenceRate LimitingQuotasThrottling

Limits

Concurrent Inference Requests account
requests
see provider documentation
Inference concurrency bounded by account capacity and credit/tier.
Concurrent Tuning Jobs account
jobs
see provider documentation
Number of simultaneous tuning jobs bounded by GPU capacity / tier.
Tuning Throughput account
steps
see provider documentation
Burst tuning scales linearly across multiple GPUs/nodes on On-Demand.
Credit / Spend Ceiling account
usd
see provider documentation
On-Demand usage is bounded by available prepaid or free credit.

Policies

Capacity-Based Limits
Throughput scales with On-Demand credit and Enterprise reserved GPU capacity rather than fixed per-minute quotas.
Backoff Strategy
Clients should implement exponential backoff with jitter and honor Retry-After on 429 responses.

Sources

Work with this as data

Every rate limit here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for rate limits

4 MCP tools reach this
  • find_rate_limitsBrowse and filter every rate limit in the catalog.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools

Call it yourself

curl for this page
This rate limit
curl "https://apis.io/api/v1/rate-limits/lamini-rate-limits"
All rate limits
curl "https://apis.io/api/v1/rate-limits?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no email required.

A second provider on the same verified email joins the account you already have.