RunPod website screenshot

RunPod

RunPod is a managed GPU cloud and serverless inference platform offering on-demand and persistent GPU Pods, autoscaling Serverless endpoints, network volumes, container templates, and a REST + GraphQL control plane for provisioning H100, H200, B200, A100, L40S, and consumer RTX GPUs. RunPod targets AI/ML developers who need flexible, per-second-billed GPU compute for training, fine-tuning, and inference workloads.

RunPod publishes 8 APIs on the APIs.io network, including Billing API, Containerregistryauth API, Docs API, and 5 more. Tagged areas include AI, Cloud, Compute, GPU, and Inference.

RunPod’s developer surface includes authentication, documentation, developer portal, signup flow, pricing, engineering blog, support, and 15 more developer resources.

54.2/100 developing ▬ flat Agent 31/100 agent aware Full breakdown ↓
scored 2026-07-28 · rubric v0.6
AccessFreeSelf serve⚡ Free to try
10 APIs 5 Features
AICloudComputeGPUInferenceMachine LearningServerless

Kin Score

Kin Score Kin Score How this is scored →
scored 2026-07-28 · rubric v0.6
Composite quality — 54.2/100 · developing
Contract Quality 13.4 / 25
Developer Ergonomics 9.6 / 20
Commercial Clarity 16.3 / 20
Operational Transparency 7.5 / 13
Governance 0.0 / 12
Discoverability 7.4 / 10
Agent readiness — 31/100 · agent aware
Machine-Readable Contract 18 / 18
Agentic Access Contract 10 / 10
MCP Server 0 / 12
Machine-Readable Auth 10 / 10
Idempotency 0 / 9
Stable Error Semantics 0 / 8
Request/Response Examples 0 / 7
Rate-Limit Signaling 7 / 7
Typed Event Surface 0 / 6
Agent Skills 0 / 5
Well-Known Catalog 0 / 4
Consent & Bot Identity 0 / 3
A2A Agent Card 0 / 8
Dry-Run / Simulate Mode 0 / 4
Improve this rating by publishing the missing artifacts — every area above can be raised, and the full rubric is at apis.io/rating/. This rating is computed from github.com/api-evangelist/runpod: open an issue to ask a question, or submit a pull request to add artifacts. Want it done for you? Prioritized profiling — $2,500 →

APIs 10

Individual APIs this provider publishes, each with its own machine-readable definition.

RunPod GraphQL API

The RunPod GraphQL API provides programmatic access to Pods, templates, and Serverless endpoints via GraphQL queries and mutations. It is the original control-plane interface an...

RunPod Serverless

RunPod Serverless provides pay-as-you-go inference endpoints with autoscaling workers, queue-based and load-balanced endpoint types, FlashBoot cold-start optimization, and per-s...

RunPod Billing API

The Billing API from RunPod — 3 operation(s) for billing.

RunPod Containerregistryauth API

The Containerregistryauth API from RunPod — 2 operation(s) for containerregistryauth.

RunPod Docs API

The Docs API from RunPod — 1 operation(s) for docs.

RunPod Endpoints API

The Endpoints API from RunPod — 3 operation(s) for endpoints.

RunPod Networkvolumes API

The Networkvolumes API from RunPod — 3 operation(s) for networkvolumes.

RunPod Openapi.json API

The Openapi.json API from RunPod — 1 operation(s) for openapi.json.

RunPod Pods API

The Pods API from RunPod — 7 operation(s) for pods.

RunPod Templates API

The Templates API from RunPod — 3 operation(s) for templates.

Scroll for all 10

Open Collections 1

Open, tool-agnostic API collections (OpenAPI-derived and Bruno).

Runpod REST API

OPEN COLLECTION

GraphQL 1

GraphQL schemas published by this provider.

RunPod GraphQL API

The RunPod GraphQL API provides programmatic access to Pods, templates, and Serverless endpoints via GraphQL queries and mutations. It is the original control-plane interface an...

GRAPHQL

Pricing Plans 1

Published pricing tiers and plan structures.

Runpod Plans Pricing

1 plans

PLANS

Rate Limits 1

Documented rate limits and quota policies.

Runpod Rate Limits

2 limits

RATE LIMITS

FinOps 1

Cost, billing, and metering signals for API financial operations.

Features 5

Notable capabilities this provider offers.

GPU Pods

Persistent on-demand GPU instances with SSH, JupyterLab, and VSCode access, billed per-second across a wide range of NVIDIA SKUs.

Serverless Endpoints

Autoscaling, queue-based inference endpoints with FlashBoot cold-start optimization and pay-per-request billing.

Network Volumes

Persistent, portable storage that can be attached to Pods and Serverless workers across datacenters.

Templates

Reusable Pod and endpoint configurations bundling container images, hardware specs, and network settings.

vLLM Quick Deploy

Pre-built Serverless workers for deploying open-source LLMs with vLLM in a single click.

Security Posture 3

Authentication, domain security, vulnerability disclosure, and trust-center signals.

Runpod Authentication

http · 1 scheme

SECURITY

Runpod Domain Security

TLSv1.3 · HSTS · DMARC

SECURITY

Runpod Trust Center

SOC 2, HIPAA, GDPR

SECURITY

Agentic Access 1

Recommended x-agentic-access execution contracts for AI agents.

Runpod Agentic Access

37 operations · 22 acting · 2 human-in-the-loop

37 operations · 22 acting

AGENTIC

Integrations 3

Pre-built integrations with other platforms and tools.

Docker

Bring-your-own container support for any Docker image on Pods and Serverless workers.

Hugging Face

Direct deployment of Hugging Face models via vLLM Quick Deploy and ready-made templates.

Pulumi

Infrastructure-as-code provisioning of RunPod resources via the official Pulumi provider.

Resources

Get Started 3

Portal, sign-up, and the first successful call

Documentation 1

Reference material describing how the API behaves

Agent Surfaces 2

MCP servers, agent skills, and machine-readable catalogs

Build 3

SDKs, sample code, and the tooling you integrate with

Access & Security 3

Authentication, authorization, and security posture

Operate 3

Status, limits, changes, and where to get help

Commercial 3

Pricing, plans, and the legal terms of use

Company 2

The organization behind the API

Other 3

Properties that don't map to a standard resource type

Source (apis.yml)

apis.yml Raw ↑
aid: runpod
name: RunPod
description: RunPod is a managed GPU cloud and serverless inference platform offering on-demand and persistent GPU Pods, autoscaling
  Serverless endpoints, network volumes, container templates, and a REST + GraphQL control plane for provisioning H100, H200,
  B200, A100, L40S, and consumer RTX GPUs. RunPod targets AI/ML developers who need flexible, per-second-billed GPU compute
  for training, fine-tuning, and inference workloads.
accessModel:
  pricing: free
  onboarding: self-serve
  trial: false
  try_now: true
  public: false
  label: Free · Self-serve signup
  confidence: high
  source:
  - plans
  - authentication
  generated: '2026-07-22'
  method: derived
image: https://kinlane-images.s3.amazonaws.com/shared/apis-json/icons/runpod.png
url: https://raw.githubusercontent.com/api-evangelist/runpod/refs/heads/main/apis.yml
created: '2026-05-23'
modified: '2026-05-23'
specificationVersion: '0.19'
type: Index
access: 3rd-Party
position: Producer
tags:
- AI
- Cloud
- Compute
- GPU
- Inference
- Machine Learning
- Serverless
apis:
- aid: runpod:graphql-api
  name: RunPod GraphQL API
  description: The RunPod GraphQL API provides programmatic access to Pods, templates, and Serverless endpoints via GraphQL
    queries and mutations. It is the original control-plane interface and is still supported alongside the REST API.
  humanURL: https://docs.runpod.io/sdks/graphql/configurations
  baseURL: https://api.runpod.io/graphql
  tags:
  - Compute
  - GPU
  - GraphQL
  - Pods
  - Serverless
  - Templates
  properties:
  - type: Documentation
    url: https://docs.runpod.io/sdks/graphql/configurations
  - url: graphql/runpod-graphql.md
    type: GraphQL
- aid: runpod:serverless
  name: RunPod Serverless
  description: RunPod Serverless provides pay-as-you-go inference endpoints with autoscaling workers, queue-based and load-balanced
    endpoint types, FlashBoot cold-start optimization, and per-second billing. Each endpoint exposes a URL that accepts request
    payloads for AI model inference and compute-intensive workloads.
  humanURL: https://docs.runpod.io/serverless/overview
  baseURL: https://api.runpod.ai/v2
  tags:
  - AI
  - Autoscaling
  - GPU
  - Inference
  - Serverless
  - Workers
  properties:
  - type: Documentation
    url: https://docs.runpod.io/serverless/overview
- aid: runpod:runpod-billing-api
  name: RunPod Billing API
  description: The Billing API from RunPod — 3 operation(s) for billing.
  humanURL: https://docs.runpod.io/api-reference/overview
  baseURL: https://rest.runpod.io/v1
  tags:
  - Billing
  properties:
  - type: OpenAPI
    url: openapi/runpod-billing-api-openapi.yml
  - type: Documentation
    url: https://docs.runpod.io/api-reference/overview
- aid: runpod:runpod-containerregistryauth-api
  name: RunPod Containerregistryauth API
  description: The Containerregistryauth API from RunPod — 2 operation(s) for containerregistryauth.
  humanURL: https://docs.runpod.io/api-reference/overview
  baseURL: https://rest.runpod.io/v1
  tags:
  - Containerregistryauth
  properties:
  - type: OpenAPI
    url: openapi/runpod-containerregistryauth-api-openapi.yml
  - type: Documentation
    url: https://docs.runpod.io/api-reference/overview
- aid: runpod:runpod-docs-api
  name: RunPod Docs API
  description: The Docs API from RunPod — 1 operation(s) for docs.
  humanURL: https://docs.runpod.io/api-reference/overview
  baseURL: https://rest.runpod.io/v1
  tags:
  - Docs
  properties:
  - type: OpenAPI
    url: openapi/runpod-docs-api-openapi.yml
  - type: Documentation
    url: https://docs.runpod.io/api-reference/overview
- aid: runpod:runpod-endpoints-api
  name: RunPod Endpoints API
  description: The Endpoints API from RunPod — 3 operation(s) for endpoints.
  humanURL: https://docs.runpod.io/api-reference/overview
  baseURL: https://rest.runpod.io/v1
  tags:
  - Endpoints
  properties:
  - type: OpenAPI
    url: openapi/runpod-endpoints-api-openapi.yml
  - type: Documentation
    url: https://docs.runpod.io/api-reference/overview
- aid: runpod:runpod-networkvolumes-api
  name: RunPod Networkvolumes API
  description: The Networkvolumes API from RunPod — 3 operation(s) for networkvolumes.
  humanURL: https://docs.runpod.io/api-reference/overview
  baseURL: https://rest.runpod.io/v1
  tags:
  - Networkvolumes
  properties:
  - type: OpenAPI
    url: openapi/runpod-networkvolumes-api-openapi.yml
  - type: Documentation
    url: https://docs.runpod.io/api-reference/overview
- aid: runpod:runpod-openapi-json-api
  name: RunPod Openapi.json API
  description: The Openapi.json API from RunPod — 1 operation(s) for openapi.json.
  humanURL: https://docs.runpod.io/api-reference/overview
  baseURL: https://rest.runpod.io/v1
  tags:
  - Openapi.json
  properties:
  - type: OpenAPI
    url: openapi/runpod-openapi-json-api-openapi.yml
  - type: Documentation
    url: https://docs.runpod.io/api-reference/overview
- aid: runpod:runpod-pods-api
  name: RunPod Pods API
  description: The Pods API from RunPod — 7 operation(s) for pods.
  humanURL: https://docs.runpod.io/api-reference/overview
  baseURL: https://rest.runpod.io/v1
  tags:
  - Pods
  properties:
  - type: OpenAPI
    url: openapi/runpod-pods-api-openapi.yml
  - type: Documentation
    url: https://docs.runpod.io/api-reference/overview
- aid: runpod:runpod-templates-api
  name: RunPod Templates API
  description: The Templates API from RunPod — 3 operation(s) for templates.
  humanURL: https://docs.runpod.io/api-reference/overview
  baseURL: https://rest.runpod.io/v1
  tags:
  - Templates
  properties:
  - type: OpenAPI
    url: openapi/runpod-templates-api-openapi.yml
  - type: Documentation
    url: https://docs.runpod.io/api-reference/overview
common:
- type: AgenticAccess
  url: agentic-access/runpod-agentic-access.yml
- type: TrustCenter
  url: security/runpod-trust-center.yml
- type: DomainSecurity
  url: security/runpod-domain-security.yml
- type: Authentication
  url: authentication/runpod-authentication.yml
- type: Website
  url: https://runpod.io
- type: Developer
  url: https://docs.runpod.io
- type: Documentation
  url: https://docs.runpod.io
- type: Portal
  url: https://console.runpod.io
- type: Signup
  url: https://www.runpod.io/console/signup
- type: Login
  url: https://www.runpod.io/console/signin
- type: Pricing
  url: https://www.runpod.io/pricing
- type: Blog
  url: https://blog.runpod.io
- type: StatusPage
  url: https://uptime.runpod.io
- type: TermsOfService
  url: https://www.runpod.io/legal/terms-of-service
- type: PrivacyPolicy
  url: https://www.runpod.io/legal/privacy-policy
- type: GitHubOrganization
  url: https://github.com/runpod
- type: Support
  url: https://www.runpod.io/contact
- type: ChangeLog
  url: https://docs.runpod.io/changelog
- name: RunPod Python SDK
  url: https://github.com/runpod/runpod-python
  type: SDKs
- name: runpodctl CLI
  url: https://github.com/runpod/runpodctl
  type: CLI
- name: RunPod Pulumi Provider
  url: https://github.com/runpod/pulumi-runpod
  type: Terraform
- type: Features
  data:
  - name: GPU Pods
    description: Persistent on-demand GPU instances with SSH, JupyterLab, and VSCode access, billed per-second across a wide
      range of NVIDIA SKUs.
  - name: Serverless Endpoints
    description: Autoscaling, queue-based inference endpoints with FlashBoot cold-start optimization and pay-per-request billing.
  - name: Network Volumes
    description: Persistent, portable storage that can be attached to Pods and Serverless workers across datacenters.
  - name: Templates
    description: Reusable Pod and endpoint configurations bundling container images, hardware specs, and network settings.
  - name: vLLM Quick Deploy
    description: Pre-built Serverless workers for deploying open-source LLMs with vLLM in a single click.
- type: Integrations
  data:
  - name: Docker
    description: Bring-your-own container support for any Docker image on Pods and Serverless workers.
  - name: Hugging Face
    description: Direct deployment of Hugging Face models via vLLM Quick Deploy and ready-made templates.
  - name: Pulumi
    description: Infrastructure-as-code provisioning of RunPod resources via the official Pulumi provider.
- type: GPUs
  data:
  - name: NVIDIA B200
    description: Blackwell-generation flagship GPU, listed at $5.98/hr.
  - name: NVIDIA H200
    description: 141GB HBM3e GPU, listed at $3.59/hr.
  - name: NVIDIA H100 SXM
    description: 80GB Hopper GPU on SXM, listed at $2.69/hr.
  - name: NVIDIA H100 PCIe
    description: 80GB Hopper GPU on PCIe, listed at $1.99/hr.
  - name: NVIDIA A100 SXM
    description: 80GB Ampere GPU on SXM, listed at $1.39/hr.
  - name: NVIDIA A100 PCIe
    description: 80GB Ampere GPU on PCIe, listed at $1.19/hr.
  - name: NVIDIA L40S
    description: 48GB Ada Lovelace inference GPU, listed at $0.79/hr.
  - name: NVIDIA RTX 4090
    description: 24GB consumer GPU, listed at $0.34/hr.
- type: LlmsText
  url: https://docs.runpod.io/llms.txt
maintainers:
- FN: Kin Lane
  email: kin@apievangelist.com
  url: https://apievangelist.com