CentML · Pricing Plans

Centml Plans Pricing

CentML uses a credit-based, pay-as-you-go pricing model where 1 CentML credit equals 1 USD. Serverless endpoints are billed by the total number of input and output tokens processed, with per-token rates that vary by model. Dedicated deployments (dedicated inference / compute endpoints) are billed by the type and duration of GPU hardware used, on a per-minute basis (commonly expressed as a per-GPU-hour rate). Enterprise plans with custom commitments are available by contacting CentML. Specific per-model token rates and per-GPU-hour rates are not reconciled in this artifact and should be confirmed on the CentML pricing page.

Centml Plans Pricing is the machine-readable pricing-plan profile for CentML on the APIs.io network, conforming to the API Commons Plans specification.

It defines 3 plans, covering usage and enterprise tiers, with named plans including Serverless Endpoints, Dedicated Deployments, Enterprise.

Tagged areas include AI, LLM, Inference, Serverless, and GPU.

3 Plans API Commons Plans
View Source
AILLMInferenceServerlessGPUPlans

Plans

Serverless Endpoints usage

OpenAI-compatible serverless inference billed by tokens processed. No infrastructure to manage; per-token rates vary by model.

Serverless Input Tokens (tokens · month) per-1M-token rate varies by model (see CentML pricing page) USD
Serverless Output Tokens (tokens · month) per-1M-token rate varies by model (see CentML pricing page) USD
Dedicated Deployments usage

Dedicated, autoscaling inference / compute endpoints billed by GPU hardware type and duration on a per-minute basis (per-GPU-hour equivalent).

Dedicated GPU Hours (gpu_hours · month) per-GPU-hour rate varies by hardware instance (A100, H100, L4, etc.) USD
Enterprise enterprise

Custom plans for larger-scale or specialized deployments, including volume commitments and dedicated support. Contact CentML sales.

Enterprise Agreement (contract · year) contact sales USD

Sources

Work with this as data

Every plan here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for plans

4 MCP tools reach this
  • find_plansBrowse and filter every plan in the catalog.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This plan
curl "https://apis.io/api/v1/plans/centml-plans-pricing"
All plans
curl "https://apis.io/api/v1/plans?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no form to fill in. Signing in shares your email address with us — we store it to create your key and to recognise you if you sign in with another provider. See our Privacy Policy and Terms.

A second provider on the same verified email joins the account you already have.