# NVIDIA NIM

**Canonical:** https://apis.io/providers/nvidia-nim/  
**Website:** https://build.nvidia.com  
**APIs profiled:** 11

NVIDIA NIM (NVIDIA Inference Microservices) is a catalog of GPU-accelerated, containerized AI inference microservices that package optimized model engines (TensorRT-LLM, vLLM, SGLang, Triton) behind industry-standard OpenAI-compatible REST APIs. NIM covers large language models, embeddings and reranking, vision-language models, speech (Riva), visual generative AI, and biology (BioNeMo) — exposed identically whether consumed from the hosted endpoint at integrate.api.nvidia.com or self-hosted via Docker containers and the Kubernetes-native NIM Operator. NIM ships with NVIDIA AI Enterprise for commercial deployment and is the inference layer underneath NVIDIA AI Blueprints, NeMo Retriever, NeMo Guardrails, and the broader NVIDIA developer stack.

## Kin Score — 71.0 / 100 (exemplar)

Scored 2026-08-20 under rubric 0.12.0. Trend: flat (+0.0 from 71.0).

| Facet | Score |
|---|---|
| Discoverability | 72.2 |
| Contract Quality | 69.2 |
| Governance | 26.5 |
| Contract Governance | 26.5 |
| Operational Transparency | 57.9 |
| Developer Ergonomics | 86.9 |
| Commercial Clarity | 92.1 |
| Access Clarity | 92.1 |

## Agent readiness — 39.7 (agent-ready)

| Dimension | Value |
|---|---|
| Spec Presence | yes |
| Agentic Access | derived |
| Reversibility Documented | no |
| MCP Server | no |
| Auth Clarity | yes |
| Idempotency | no |
| Error Semantics | documented |
| OpenAPI Examples | partial |
| Rate Limit Signal | documented |
| Event Surface Described | no |
| Agent Skills | yes |
| Well Known Catalog | no |
| Consent Identity | no |
| Agent Card | no |
| Dry Run Mode | no |

## Access

Freemium · Self-serve signup — onboarding: self-serve, pricing: freemium, trial: no (confidence: high).

## APIs (11)

- **NVIDIA NIM Completions API** — Legacy OpenAI-compatible text completion endpoint (/v1/completions) for non-chat foundation models served by NIM. Accepts a raw prompt and returns generated text with the same s...
- **NVIDIA NIM Embeddings API** — OpenAI-compatible embeddings endpoint (/v1/embeddings) backed by NVIDIA NeMo Retriever text embedding models including NV-Embed, NV-EmbedQA-E5, llama-3.2-nv-embedqa-1b, and BAAI...
- **NVIDIA NIM Reranking API** — NeMo Retriever cross-encoder reranking endpoint (/v1/ranking) for scoring candidate passages against a query. Improves retrieval relevance on RAG pipelines and supports the llam...
- **NVIDIA NIM Models API** — OpenAI-compatible model catalog endpoint (/v1/models) returning the list of models served by the NIM endpoint or container. Each entry includes id, owned_by, and created timesta...
- **NVIDIA NIM Vision Language Models API** — Vision-language model inference through the standard /v1/chat/completions surface with image inputs (base64 or URL) in the messages payload. Supports NVIDIA NeVA, microsoft/kosm...
- **NVIDIA NIM Health API** — Liveness, readiness, and startup probes exposed by self-hosted NIM containers (/v1/health/live, /v1/health/ready) and a Prometheus /v1/metrics scrape endpoint for GPU utilizatio...
- **NVIDIA NIM Biology (BioNeMo) API** — BioNeMo NIMs for protein structure prediction (AlphaFold2, ESMFold, OpenFold), protein generation (ProtGPT2, RFDiffusion), molecular property prediction (MolMIM), small molecule...
- **NVIDIA NIM ASR API** — Automatic speech recognition (speech-to-text)
- **NVIDIA NIM Chat API** — OpenAI-compatible chat completion operations
- **NVIDIA NIM Images API** — Text-to-image and image-to-image generation
- **NVIDIA NIM TTS API** — Text-to-speech synthesis

## MCP servers (1)

- **NVIDIA NIM MCP Server**

## Agentic access (1)

- **Nvidia Nim Agentic Access** — 16 operations · 11 acting

## Security (3)

- **Nvidia Nim Authentication** — http · 1 scheme
- **Nvidia Nim Domain Security** — TLSv1.3 · HSTS · DMARC
- **Nvidia Nim Vulnerability Disclosure** — disclosure policy published

## Plans (1)

- **Nvidia Nim Plans Pricing**

## Tags

Artificial Intelligence, Inference, Microservices, LLM, Foundation Models, GPU, Kubernetes, NVIDIA, OpenAI-Compatible

---

Profiled by [API Evangelist](https://apievangelist.com) and published on [APIs.io](https://apis.io/providers/nvidia-nim/). Scores are computed from the provider's own public artifacts under a published rubric.
