# Hypura Ollama-Compatible Inference API

**Canonical:** https://apis.io/apis/community-labs/hypura/  
**Provider:** Community Labs — https://apis.io/providers/community-labs/  
**Base URL:** http://127.0.0.1:8080  
**Documentation:** https://github.com/t8/hypura

Hypura Ollama-Compatible Inference API is published by [Community Labs](https://apis.io/providers/community-labs/) on the [APIs.io](https://apis.io/) network. Tagged areas include LLM Inference, Local AI, Ollama Compatible, and Apple Silicon. The published artifact set on APIs.io includes API documentation, an API reference, and a getting-started guide.

Hypura is a storage-tier-aware LLM inference scheduler for Apple Silicon that places model tensors across GPU, RAM, and NVMe tiers so models larger than physical memory can run. Running `hypura serve <model.gguf>` exposes a local Ollama-compatible HTTP API, making it a drop-in replacement for tooling that talks to Ollama. The server is local-first; there is no hosted endpoint.

## Machine-readable artifacts (6)

- **SourceCode** — https://github.com/t8/hypura
- **Documentation** — https://github.com/t8/hypura#readme
- **APIReference** — https://github.com/t8/hypura#endpoints
- **GettingStarted** — https://github.com/t8/hypura#quick-start
- **Conformance** — https://raw.githubusercontent.com/api-evangelist/community-labs/refs/heads/main/conformance/community-labs-conformance.yml
- **APIsJSON** — https://raw.githubusercontent.com/api-evangelist/community-labs/refs/heads/main/apis.yml

## Tags

LLM Inference, Local AI, Ollama Compatible, Apple Silicon

---

Profiled by [API Evangelist](https://apievangelist.com) and published on [APIs.io](https://apis.io/apis/community-labs/hypura/). The API's provider profile, Kin Score and agent-readiness rating are at https://apis.io/providers/community-labs/.
