vLLM · Pricing Plans

Vllm Plans Pricing

vLLM is free open-source software (Apache 2.0). The project does not sell hosting. Cost is incurred entirely on your own GPU infrastructure (cloud or on-prem). Several third parties (RunPod, Modal, Anyscale, Baseten, etc.) offer managed vLLM hosting, billed by them — not by the vLLM project.

Vllm Plans Pricing is the machine-readable pricing-plan profile for vLLM on the APIs.io network, conforming to the API Commons Plans specification.

It defines 1 plan, covering free tiers, with named plans including Self-Hosted (Apache 2.0).

Tagged areas include LLM, Inference, Open Source, GPU, and OpenAI Compatible.

1 Plans API Commons Plans
View Source
LLMInferenceOpen SourceGPUOpenAI CompatibleSelf-HostedPlans

Plans

Self-Hosted (Apache 2.0) free

Run vLLM on your own GPU infrastructure under Apache 2.0.

Self-Host (deployment · lifetime) 0 USD

Sources

Work with this as data

Every plan here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for plans

4 MCP tools reach this
  • find_plansBrowse and filter every plan in the catalog.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools

Call it yourself

curl for this page
This plan
curl "https://apis.io/api/v1/plans/vllm-plans-pricing"
All plans
curl "https://apis.io/api/v1/plans?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no email required.

A second provider on the same verified email joins the account you already have.