Nebius website screenshot

Nebius

Nebius is an AI-focused cloud platform spun out of Yandex, offering NVIDIA GPU virtual machines and clusters (GB300, GB200, B300, B200, H200, H100) connected over InfiniBand, managed Kubernetes and Slurm (Soperator), S3-compatible storage, managed PostgreSQL, container registry, MLflow, JupyterLab, vLLM, and the Token Factory inference platform. Nebius exposes a gRPC control plane API, a Terraform provider, a nebius CLI, and Go and TypeScript SDKs.

Nebius publishes 6 APIs on the APIs.io network. Tagged areas include AI, Cloud, Compute, GPU, and HPC.

Nebius’ developer surface includes documentation, developer portal, pricing, engineering blog, support, and 14 more developer resources.

28.1/100 thin ▬ flat Agent 3/100 human only Full breakdown ↓
scored 2026-08-05 · rubric v0.9.1
AccessFree
6 APIs 7 Features
AICloudComputeGPUHPCInferenceKubernetesMachine LearningStorage

Kin Score

Kin Score Kin Score How this is scored →
scored 2026-08-05 · rubric v0.9.1
Composite quality — 28.1/100 · thin
Contract Quality 0.0 / 25
Developer Ergonomics 6.1 / 20
Commercial Clarity 12.1 / 20
Operational Transparency 3.4 / 13
Governance 0.0 / 12
Discoverability 6.5 / 10
Agent readiness — 3/100 · human only
Machine-Readable Contract 0 / 18
Agentic Access Contract 0 / 10
MCP Server 0 / 12
Machine-Readable Auth 0 / 10
Idempotency 0 / 9
Stable Error Semantics 0 / 8
Request/Response Examples 0 / 7
Rate-Limit Signaling 7 / 7
Typed Event Surface 0 / 6
Agent Skills 0 / 5
Well-Known Catalog 0 / 4
Consent & Bot Identity 0 / 3
A2A Agent Card 0 / 8
Dry-Run / Simulate Mode 0 / 4
Improve this rating by publishing the missing artifacts — every area above can be raised, and the full rubric is at apis.io/rating/. This rating is computed from github.com/api-evangelist/nebius: open an issue to ask a question, or submit a pull request to add artifacts. Want it done for you? Prioritized profiling — $2,500 →

APIs 6

Individual APIs this provider publishes, each with its own machine-readable definition.

Nebius Compute API

The Nebius Compute API provisions and manages virtual machines and GPU clusters with NVIDIA GPUs and InfiniBand interconnect for ML and AI workloads. Exposed over gRPC and acces...

Nebius Managed Kubernetes API

The Managed Kubernetes API provisions Kubernetes clusters with GPU and InfiniBand support for distributed training and inference workloads.

Nebius Storage API

The Nebius Storage API exposes AWS S3-compatible object storage for ML/AI datasets and model artifacts.

Nebius IAM API

The Nebius Identity and Access Management API controls users, service accounts, projects, and resource-level access policies.

Nebius Managed Applications API

The Managed Applications API deploys and manages JupyterLab, vLLM, Open WebUI, MLflow, and other ready-made apps on Nebius infrastructure.

Nebius Token Factory

Nebius Token Factory is the AI model inference platform offering OpenAI-compatible endpoints for serving open-source LLMs on Nebius GPU infrastructure.

Pricing Plans 1

Published pricing tiers and plan structures.

Nebius Plans Pricing

1 plans

PLANS

Rate Limits 1

Documented rate limits and quota policies.

Nebius Rate Limits

2 limits

RATE LIMITS

FinOps 1

Cost, billing, and metering signals for API financial operations.

Features 7

Notable capabilities this provider offers.

GPU Compute

Virtual machines and clusters with NVIDIA GB300, GB200, B300, B200, H200, and H100 GPUs.

InfiniBand Networking

High-bandwidth InfiniBand interconnect for large-scale distributed training.

Managed Kubernetes

Kubernetes clusters with GPU and InfiniBand support.

Slurm via Soperator

Slurm workload manager running on Kubernetes via the open-source Soperator project.

S3-Compatible Storage

Object storage optimized for ML datasets and model artifacts.

Managed Applications

One-click JupyterLab, vLLM, Open WebUI, and MLflow deployments.

Token Factory

OpenAI-compatible inference endpoints for open-source LLMs.

Scroll for all 7

Security Posture 2

Authentication, domain security, vulnerability disclosure, and trust-center signals.

Nebius Domain Security

TLSv1.3 · HSTS · DMARC

SECURITY

Nebius Vulnerability Disclosure

security.txt · contact published

SECURITY

Integrations 5

Pre-built integrations with other platforms and tools.

Terraform

Official Terraform provider for Nebius infrastructure.

Kubernetes

Standard Kubernetes API across managed clusters.

Slurm

Slurm workload manager via the open-source Soperator project.

MLflow

Managed MLflow for experiment tracking.

PostgreSQL

Managed PostgreSQL database clusters.

Resources

Get Started 1

Portal, sign-up, and the first successful call

Documentation 1

Reference material describing how the API behaves

Agent Surfaces 1

MCP servers, agent skills, and machine-readable catalogs

Build 5

SDKs, sample code, and the tooling you integrate with

Access & Security 2

Authentication, authorization, and security posture

Operate 1

Status, limits, changes, and where to get help

Commercial 3

Pricing, plans, and the legal terms of use

Company 3

The organization behind the API

Other 3

Properties that don't map to a standard resource type

Source (apis.yml)

apis.yml Raw ↑
aid: nebius
name: Nebius
description: Nebius is an AI-focused cloud platform spun out of Yandex, offering NVIDIA GPU virtual machines and clusters
  (GB300, GB200, B300, B200, H200, H100) connected over InfiniBand, managed Kubernetes and Slurm (Soperator), S3-compatible
  storage, managed PostgreSQL, container registry, MLflow, JupyterLab, vLLM, and the Token Factory inference platform. Nebius
  exposes a gRPC control plane API, a Terraform provider, a nebius CLI, and Go and TypeScript SDKs.
accessModel:
  pricing: free
  onboarding: unknown
  trial: false
  try_now: false
  public: false
  label: Free
  confidence: medium
  source:
  - plans
  generated: '2026-07-22'
  method: derived
image: https://kinlane-images.s3.amazonaws.com/shared/apis-json/icons/nebius.png
url: https://raw.githubusercontent.com/api-evangelist/nebius/refs/heads/main/apis.yml
created: '2026-05-23'
modified: '2026-05-23'
specificationVersion: '0.19'
type: Index
access: 3rd-Party
position: Producer
tags:
- AI
- Cloud
- Compute
- GPU
- HPC
- Inference
- Kubernetes
- Machine Learning
- Storage
apis:
- aid: nebius:compute-api
  name: Nebius Compute API
  description: The Nebius Compute API provisions and manages virtual machines and GPU clusters with NVIDIA GPUs and InfiniBand
    interconnect for ML and AI workloads. Exposed over gRPC and accessed through the nebius CLI, Terraform provider, and language
    SDKs.
  humanURL: https://docs.nebius.com/compute
  tags:
  - Compute
  - GPU
  - InfiniBand
  - Virtual Machines
  properties:
  - type: Documentation
    url: https://docs.nebius.com/compute
- aid: nebius:kubernetes-api
  name: Nebius Managed Kubernetes API
  description: The Managed Kubernetes API provisions Kubernetes clusters with GPU and InfiniBand support for distributed training
    and inference workloads.
  humanURL: https://docs.nebius.com/kubernetes
  tags:
  - Clusters
  - Containers
  - GPU
  - Kubernetes
  properties:
  - type: Documentation
    url: https://docs.nebius.com/kubernetes
- aid: nebius:storage-api
  name: Nebius Storage API
  description: The Nebius Storage API exposes AWS S3-compatible object storage for ML/AI datasets and model artifacts.
  humanURL: https://docs.nebius.com/storage
  tags:
  - Datasets
  - Object Storage
  - S3
  - Storage
  properties:
  - type: Documentation
    url: https://docs.nebius.com/storage
- aid: nebius:iam-api
  name: Nebius IAM API
  description: The Nebius Identity and Access Management API controls users, service accounts, projects, and resource-level
    access policies.
  humanURL: https://docs.nebius.com/iam
  tags:
  - Access Control
  - IAM
  - Identity
  - Service Accounts
  properties:
  - type: Documentation
    url: https://docs.nebius.com/iam
- aid: nebius:mk8s-applications-api
  name: Nebius Managed Applications API
  description: The Managed Applications API deploys and manages JupyterLab, vLLM, Open WebUI, MLflow, and other ready-made
    apps on Nebius infrastructure.
  humanURL: https://docs.nebius.com/applications
  tags:
  - Applications
  - JupyterLab
  - MLflow
  - vLLM
  properties:
  - type: Documentation
    url: https://docs.nebius.com/applications
- aid: nebius:token-factory
  name: Nebius Token Factory
  description: Nebius Token Factory is the AI model inference platform offering OpenAI-compatible endpoints for serving open-source
    LLMs on Nebius GPU infrastructure.
  humanURL: https://docs.tokenfactory.nebius.com/quickstart
  tags:
  - AI
  - Inference
  - LLMs
  - OpenAI Compatible
  - Serverless
  properties:
  - type: Documentation
    url: https://docs.tokenfactory.nebius.com/quickstart
  - type: Portal
    url: https://tokenfactory.nebius.com
common:
- type: VulnerabilityDisclosure
  url: security/nebius-vulnerability-disclosure.yml
- type: DomainSecurity
  url: security/nebius-domain-security.yml
- type: Website
  url: https://nebius.com
- type: Developer
  url: https://docs.nebius.com
- type: Documentation
  url: https://docs.nebius.com
- type: Portal
  url: https://console.nebius.com
- type: Pricing
  url: https://nebius.com/prices
- type: Blog
  url: https://nebius.com/blog
- type: GitHubOrganization
  url: https://github.com/nebius
- type: TermsOfService
  url: https://nebius.com/legal/terms-of-service
- type: PrivacyPolicy
  url: https://nebius.com/legal/privacy-policy
- type: Support
  url: https://nebius.com/contact
- type: LinkedIn
  url: https://www.linkedin.com/company/nebius
- name: Nebius Go SDK
  url: https://github.com/nebius/gosdk
  type: SDKs
- name: Nebius JavaScript SDK
  url: https://github.com/nebius/js-sdk
  type: SDKs
- name: Nebius Terraform Provider
  url: https://github.com/nebius/terraform-provider-nebius
  type: Terraform
- name: Nebius Solution Library
  url: https://github.com/nebius/nebius-solution-library
  type: Samples
- name: Soperator
  url: https://github.com/nebius/soperator
  type: Samples
- type: Features
  data:
  - name: GPU Compute
    description: Virtual machines and clusters with NVIDIA GB300, GB200, B300, B200, H200, and H100 GPUs.
  - name: InfiniBand Networking
    description: High-bandwidth InfiniBand interconnect for large-scale distributed training.
  - name: Managed Kubernetes
    description: Kubernetes clusters with GPU and InfiniBand support.
  - name: Slurm via Soperator
    description: Slurm workload manager running on Kubernetes via the open-source Soperator project.
  - name: S3-Compatible Storage
    description: Object storage optimized for ML datasets and model artifacts.
  - name: Managed Applications
    description: One-click JupyterLab, vLLM, Open WebUI, and MLflow deployments.
  - name: Token Factory
    description: OpenAI-compatible inference endpoints for open-source LLMs.
- type: GPUs
  data:
  - name: NVIDIA GB300
    description: Grace Blackwell Ultra superchip.
  - name: NVIDIA GB200
    description: Grace Blackwell superchip.
  - name: NVIDIA B300
    description: Blackwell Ultra GPU.
  - name: NVIDIA B200
    description: Blackwell GPU.
  - name: NVIDIA H200
    description: 141GB HBM3e Hopper GPU.
  - name: NVIDIA H100
    description: 80GB Hopper GPU.
- type: Integrations
  data:
  - name: Terraform
    description: Official Terraform provider for Nebius infrastructure.
  - name: Kubernetes
    description: Standard Kubernetes API across managed clusters.
  - name: Slurm
    description: Slurm workload manager via the open-source Soperator project.
  - name: MLflow
    description: Managed MLflow for experiment tracking.
  - name: PostgreSQL
    description: Managed PostgreSQL database clusters.
- type: LlmsText
  url: https://docs.nebius.com/llms.txt
maintainers:
- FN: Kin Lane
  email: kin@apievangelist.com
  url: https://apievangelist.com