Triton Inference Server · Schema
Triton Inference Request
An inference request submitted to NVIDIA Triton Inference Server following the KServe V2 inference protocol. Contains input tensors, optional output specifications, and inference parameters for sequence handling, priority, and timeout control.
Artificial IntelligenceDeep LearningInferenceMachine-LearningModel ServingNVIDIAOpen-Source
Properties
JSON Schema
Work with this as data
Every JSON Schema here is available over the APIs.io API and to AI agents over MCP.
MCP server
One button, every client — Claude, Cursor, VS Code and the rest.
https://apis.io/mcp
Tools for schemas
4 MCP tools reach this
find_json_schemasBrowse and filter every JSON Schema in the catalog.apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.resolveTurn a domain, URL or GitHub org into the provider it belongs to.find_cohortsEvery scored population of providers in the catalog.
Call it yourself
curl for this page
This JSON Schema
curl "https://apis.io/api/v1/json-schemas/triton-inference-request"
All schemas
curl "https://apis.io/api/v1/json-schemas?limit=25"
Discovery needs no key. Ratings and market analysis are Pro.