Nebius
Nebius is an AI-focused cloud platform spun out of Yandex, offering NVIDIA GPU virtual machines and clusters (GB300, GB200, B300, B200, H200, H100) connected over InfiniBand, managed Kubernetes and Slurm (Soperator), S3-compatible storage, managed PostgreSQL, container registry, MLflow, JupyterLab, vLLM, and the Token Factory inference platform. Nebius exposes a gRPC control plane API, a Terraform provider, a nebius CLI, and Go and TypeScript SDKs.
Nebius publishes 6 APIs on the APIs.io network. Tagged areas include AI, Cloud, Compute, GPU, and HPC.
Nebius’ developer surface includes documentation, developer portal, pricing, engineering blog, support, and 14 more developer resources.
Kin Score
APIs 6
Individual APIs this provider publishes, each with its own machine-readable definition.
Nebius Compute API
The Nebius Compute API provisions and manages virtual machines and GPU clusters with NVIDIA GPUs and InfiniBand interconnect for ML and AI workloads. Exposed over gRPC and acces...
Nebius Managed Kubernetes API
The Managed Kubernetes API provisions Kubernetes clusters with GPU and InfiniBand support for distributed training and inference workloads.
Nebius Storage API
The Nebius Storage API exposes AWS S3-compatible object storage for ML/AI datasets and model artifacts.
Nebius IAM API
The Nebius Identity and Access Management API controls users, service accounts, projects, and resource-level access policies.
Nebius Managed Applications API
The Managed Applications API deploys and manages JupyterLab, vLLM, Open WebUI, MLflow, and other ready-made apps on Nebius infrastructure.
Nebius Token Factory
Nebius Token Factory is the AI model inference platform offering OpenAI-compatible endpoints for serving open-source LLMs on Nebius GPU infrastructure.
Pricing Plans 1
Published pricing tiers and plan structures.
Nebius Plans Pricing
PLANSRate Limits 1
Documented rate limits and quota policies.
Nebius Rate Limits
RATE LIMITSFinOps 1
Cost, billing, and metering signals for API financial operations.
Nebius Finops
FINOPSFeatures 7
Notable capabilities this provider offers.
GPU Compute
Virtual machines and clusters with NVIDIA GB300, GB200, B300, B200, H200, and H100 GPUs.
InfiniBand Networking
High-bandwidth InfiniBand interconnect for large-scale distributed training.
Managed Kubernetes
Kubernetes clusters with GPU and InfiniBand support.
Slurm via Soperator
Slurm workload manager running on Kubernetes via the open-source Soperator project.
S3-Compatible Storage
Object storage optimized for ML datasets and model artifacts.
Managed Applications
One-click JupyterLab, vLLM, Open WebUI, and MLflow deployments.
Token Factory
OpenAI-compatible inference endpoints for open-source LLMs.
Scroll for all 7
Security Posture 2
Authentication, domain security, vulnerability disclosure, and trust-center signals.
Integrations 5
Pre-built integrations with other platforms and tools.
Terraform
Official Terraform provider for Nebius infrastructure.
Kubernetes
Standard Kubernetes API across managed clusters.
Slurm
Slurm workload manager via the open-source Soperator project.
MLflow
Managed MLflow for experiment tracking.
PostgreSQL
Managed PostgreSQL database clusters.
Resources
Get Started 1
Portal, sign-up, and the first successful call
Documentation 1
Reference material describing how the API behaves
Agent Surfaces 1
MCP servers, agent skills, and machine-readable catalogs
Build 5
SDKs, sample code, and the tooling you integrate with
Access & Security 2
Authentication, authorization, and security posture
Operate 1
Status, limits, changes, and where to get help
Commercial 3
Pricing, plans, and the legal terms of use
Company 3
The organization behind the API
Other 3
Properties that don't map to a standard resource type