Crusoe Managed Inference API
OpenAI-compatible inference API from the Crusoe Intelligence Foundry. Send chat/completions and embeddings requests to Crusoe-hosted open models (DeepSeek, Llama, Gemma, GLM, Kimi, Nemotron and others) without managing GPU infrastructure, or run reserved-capacity self-serve deployments and LoRA-based serverless fine-tuning jobs. Authenticated with an Inference API key issued from the Crusoe Cloud console. The endpoint requires credentials for every request, including discovery paths, so no anonymous machine-readable contract is published for it.
Documentation
Documentation
https://docs.crusoecloud.com/serverless-inference/overview
GettingStarted
https://docs.crusoecloud.com/quickstart/getting-started-with-serverless-inference
Authentication
https://raw.githubusercontent.com/api-evangelist/crusoe/refs/heads/main/authentication/crusoe-authentication.yml