# Apache Hudi

**Canonical:** https://apis.io/providers/apache-hudi/  
**APIs profiled:** 3

Apache Hudi is a data lake platform that provides incremental data processing primitives including upserts and incremental queries. It manages storage of large analytical datasets on distributed file systems with ACID transactions, timeline-based versioning, and integrations for Spark, Flink, and Hive.

## Kin Score — 33.9 / 100 (thin)

Scored 2026-08-20 under rubric 0.12.0. Trend: flat (+0.0 from 33.9).

| Facet | Score |
|---|---|
| Discoverability | 64.8 |
| Contract Quality | 52.4 |
| Governance | 25.0 |
| Contract Governance | 25.0 |
| Operational Transparency | 26.3 |
| Developer Ergonomics | 23.8 |
| Commercial Clarity | 15.8 |
| Access Clarity | 15.8 |

## Agent readiness — 20.5 (agent-aware)

| Dimension | Value |
|---|---|
| Spec Presence | yes |
| Agentic Access | derived |
| Reversibility Documented | no |
| MCP Server | no |
| Auth Clarity | no |
| Idempotency | no |
| Error Semantics | no |
| OpenAPI Examples | no |
| Rate Limit Signal | documented |
| Event Surface Described | no |
| Agent Skills | no |
| Well Known Catalog | no |
| Consent Identity | no |
| Agent Card | no |
| Dry Run Mode | no |

## Access

Freemium — onboarding: unknown, pricing: freemium, trial: no (confidence: medium).

## APIs (3)

- **Apache Hudi Java API** — Java API for writing Hudi tables with upserts, inserts, and deletes, plus timeline management, compaction, and Spark/Flink DataSource integration APIs.
- **Apache Hudi Tables API** — Hudi table management operations
- **Apache Hudi Timeline API** — Hudi timeline and commit operations

## Agentic access (1)

- **Apache Hudi Agentic Access** — 5 operations · 2 acting

## Security (2)

- **Apache Hudi Domain Security** — TLSv1.3 · HSTS · DMARC
- **Apache Hudi Vulnerability Disclosure** — security.txt · contact published

## Plans (1)

- **Apache Hudi Plans Pricing**

## Use cases (5)

- **CDC Pipeline Ingestion** — Ingest change data capture (CDC) events from databases into data lake tables with upsert support.
- **Streaming Data Lake** — Build near-real-time data lake pipelines with Spark Structured Streaming or Flink.
- **Data Lake Maintenance** — Manage storage costs with automated cleaning, compaction, and clustering of Hudi tables.
- **Incremental ETL** — Build incremental ETL pipelines that process only changed data since the last run.
- **Regulatory Data Retention** — Implement GDPR right-to-erasure by deleting records from Hudi tables with delete operations.

## Tags

ACID, Apache, Big Data, Data Lake, Incremental Processing, Lakehouse, Open-Source

---

Profiled by [API Evangelist](https://apievangelist.com) and published on [APIs.io](https://apis.io/providers/apache-hudi/). Scores are computed from the provider's own public artifacts under a published rubric.
