Mango LLMBoost Inference Server API

LLMBoost is MangoBoost's enterprise LLM inference server. It serves the OpenAI REST API on /v1 so an existing OpenAI client migrates with a base-URL change: POST /v1/chat/completions, POST /v1/completions, POST /v1/embeddings, POST /v1/responses, POST /v1/audio/transcriptions and /translations, GET /v1/models, plus GET /health and GET /metrics (Prometheus). Streaming, JSON-schema structured output, tool/function calling and multimodal image input are supported. The server is self-hosted on the customer's own AMD Instinct GPUs — there is no MangoBoost-hosted endpoint — and is started with `lbh serve ` or `llmboost serve ` inside the LLMBoost container.

Work with this as data

Every API here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for apis

7 MCP tools reach this
  • find_apisBrowse and filter every API in the catalog.
  • get_api_artifactsOne API's artifacts, grouped by type.
  • get_openapiThe primary OpenAPI for this API.
  • find_similar_apisAPIs that look like this one.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This API
curl "https://apis.io/api/v1/apis/llmboost-inference"
All apis
curl "https://apis.io/api/v1/apis?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no form to fill in. Signing in shares your email address with us — we store it to create your key and to recognise you if you sign in with another provider. See our Privacy Policy and Terms.

A second provider on the same verified email joins the account you already have.

API entry from apis.yml

apis.yml Raw ↑
aid: mangoboost:llmboost-inference
name: Mango LLMBoost Inference Server API
description: 'LLMBoost is MangoBoost''s enterprise LLM inference server. It serves the OpenAI REST API
  on /v1 so an existing OpenAI client migrates with a base-URL change: POST /v1/chat/completions, POST
  /v1/completions, POST /v1/embeddings, POST /v1/responses, POST /v1/audio/transcriptions and /translations,
  GET /v1/models, plus GET /health and GET /metrics (Prometheus). Streaming, JSON-schema structured output,
  tool/function calling and multimodal image input are supported. The server is self-hosted on the customer''s
  own AMD Instinct GPUs — there is no MangoBoost-hosted endpoint — and is started with `lbh serve <Repo/Model>`
  or `llmboost serve <model>` inside the LLMBoost container.'
humanURL: https://llmboost.mangoboost.io/docs/features/openai-api
baseURL: http://localhost:8000/v1
x-deployment: self-hosted
x-default-port: 8000
tags:
- Inference
- Artificial Intelligence
- LLM
- OpenAI-Compatible
- Self-Hosted
tags_raw:
- Inference
- Artificial Intelligence
- LLM
- OpenAI Compatible
- Self Hosted
properties:
- type: Documentation
  url: https://llmboost.mangoboost.io/docs/
- type: GettingStarted
  url: https://llmboost.mangoboost.io/docs/quickstart
- type: APIReference
  url: https://llmboost.mangoboost.io/docs/features/openai-api