Vllm Plans Pricing
vLLM is free open-source software (Apache 2.0). The project does not sell hosting. Cost is incurred entirely on your own GPU infrastructure (cloud or on-prem). Several third parties (RunPod, Modal, Anyscale, Baseten, etc.) offer managed vLLM hosting, billed by them — not by the vLLM project.
Vllm Plans Pricing is the machine-readable pricing-plan profile for vLLM on the APIs.io network, conforming to the API Commons Plans specification.
It defines 1 plan, covering free tiers, with named plans including Self-Hosted (Apache 2.0).
Tagged areas include LLM, Inference, Open Source, GPU, and OpenAI Compatible.
Plans
Run vLLM on your own GPU infrastructure under Apache 2.0.
- pip install vllm
- vllm serve
- Bring your own GPU
Sources
Work with this as data
Every plan here is available over the APIs.io API and to AI agents over MCP.
MCP server
One button, every client — Claude, Cursor, VS Code and the rest.
https://apis.io/mcp
Tools for plans
4 MCP tools reach this
find_plansBrowse and filter every plan in the catalog.apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.resolveTurn a domain, URL or GitHub org into the provider it belongs to.find_cohortsEvery scored population of providers in the catalog.
Call it yourself
curl for this page
curl "https://apis.io/api/v1/plans/vllm-plans-pricing"
curl "https://apis.io/api/v1/plans?limit=25"
Discovery needs no key. Ratings and market analysis are Pro.