Centml Finops
FinOps view of CentML spend. CentML bills on a credit-based model where 1 credit equals 1 USD. Serverless inference is metered by input and output tokens with per-token rates that vary by model. Dedicated deployments are metered by GPU hardware type and duration on a per-minute basis (per-GPU-hour equivalent). Exact per-model and per-GPU-hour rates are not reconciled here.
Centml Finops is the FinOps profile for CentML on the APIs.io network, aligned with the FinOps Foundation Framework.
It defines 4 billable meters, billed in USD, on a monthly cycle, and pricing category usage-based.
The profile maps 8 FOCUS columns for cost-allocation reporting.
Tagged areas include AI, LLM, Inference, Serverless, and GPU.
Framework Alignment
Charge Categories
FOCUS Columns
Meters
Sources
- https://centml.ai/pricing/
- https://docs.centml.ai/apps/serverless
- https://focus.finops.org/focus-specification/v1-3/
Work with this as data
Every finops artifact here is available over the APIs.io API and to AI agents over MCP.