Apache Druid
Apache Druid is a high-performance, real-time analytics database governed by the Apache Software Foundation, designed for fast slice-and-dice OLAP queries on event-time data. It features a distributed, column-oriented storage engine with automatic rollup, supports both streaming (Kafka, Kinesis) and batch (S3, HDFS, local) data ingestion, and provides a SQL query interface plus a native JSON query API via REST. Druid is optimized for sub-second queries at petabyte scale with high concurrency.
Apache Druid publishes 1 API on the APIs.io network: Druid API. Tagged areas include Analytics, Apache, Database, Kafka, and OLAP.
The Apache Druid catalog on APIs.io includes 1 JSON-LD context and 1 Spectral governance ruleset.
Apache Druid’s developer surface includes developer portal, documentation, getting-started guide, engineering blog, Stack Overflow tag, and 6 more developer resources.
Kin Score
APIs 1
Individual APIs this provider publishes, each with its own machine-readable definition.
Apache Druid Druid API
The Druid API from Apache Druid — 10 operation(s) for druid.
Open Collections 1
Open, tool-agnostic API collections (OpenAPI-derived and Bruno).
Apache Druid REST API
OPEN COLLECTIONPricing Plans 1
Published pricing tiers and plan structures.
Rate Limits 1
Documented rate limits and quota policies.
Apache Druid Rate Limits
RATE LIMITSFinOps 1
Cost, billing, and metering signals for API financial operations.
Apache Druid Finops
FINOPSFeatures 8
Notable capabilities this provider offers.
Sub-Second OLAP Queries
Columnar storage with bitmap indexes, dictionary encoding, and pre-aggregation (rollup) enables sub-second queries on billions of events.
Druid SQL API
REST endpoint for submitting standard SQL queries with ANSI SQL support, time-based filtering, and streaming response options.
Native JSON Query API
Druid-native query format (Timeseries, TopN, GroupBy, Scan, Search) for maximum control and performance.
Streaming Ingestion
Real-time data ingestion from Apache Kafka and Amazon Kinesis with supervisor-managed offset tracking and exactly-once semantics.
Batch Ingestion
Parallel batch indexing tasks from local files, S3, GCS, HDFS, and other external storage systems.
Automatic Rollup
Pre-aggregates metrics at ingestion time to reduce storage and query time, configurable per datasource.
Time-Based Partitioning
All data is partitioned by time interval (segments), enabling efficient time-range query pruning.
Multi-Tenancy
Query isolation and resource management via query lanes, scheduler priorities, and row-level access control.
Scroll for all 8
Semantic Vocabularies 1
JSON-LD contexts and semantic vocabularies used across these APIs.
Apache Druid Context
JSON-LDSpectral Rules 1
Spectral governance rulesets for linting and validating these APIs.
Apache Druid API Rules
SPECTRALJSON Schema 4
Standalone JSON Schema definitions for this provider's data models.
IngestionTask
JSON SCHEMASqlQueryRequest
JSON SCHEMASqlQueryResponse
JSON SCHEMASupervisor
JSON SCHEMAJSON Structure 4
JSON Structure definitions describing this provider's data shapes.
Apache Druid Ingestion Task Structure
JSON STRUCTUREApache Druid Sql Query Request Structure
JSON STRUCTUREApache Druid Sql Query Response Structure
JSON STRUCTUREApache Druid Supervisor Structure
JSON STRUCTUREExamples 4
Example request and response payloads for these APIs.
Security Posture 2
Authentication, domain security, vulnerability disclosure, and trust-center signals.
Agentic Access 1
Recommended x-agentic-access execution contracts for AI agents.
Use Cases 5
What developers build with this provider.
Real-Time Event Analytics
Analyze click streams, IoT events, application logs, and user behavior data with sub-second query latency.
Business Intelligence Dashboards
Power interactive BI dashboards with high-concurrency low-latency queries backed by Druid's columnar engine.
Network and Security Monitoring
Ingest and analyze network flow data and security events in real time for threat detection and capacity planning.
Ad Tech Analytics
Process advertising impression, click, and conversion events at high volume with real-time aggregation.
Operational Analytics
Monitor application performance metrics and operational data with drilldown and filtering capabilities.
Integrations 7
Pre-built integrations with other platforms and tools.
Apache Kafka
KafkaSupervisor for real-time continuous ingestion from Kafka topics into Druid datasources.
Amazon Kinesis
KinesisSupervisor for real-time data ingestion from AWS Kinesis data streams.
Apache Hadoop / HDFS
Native Hadoop batch indexing task for bulk loading data from HDFS or MapReduce job outputs.
Amazon S3 / GCS
Batch and streaming ingestion from object storage (S3, GCS, Azure Blob) using index tasks.
Apache Hive
Druid-Hive integration for querying Druid datasources from HiveQL and performing joins.
Kubernetes
Official Kubernetes operator for deploying and managing Druid clusters on Kubernetes.
Imply (Commercial)
Imply provides a commercial managed Druid service with additional features and enterprise support.
Scroll for all 7
Resources
Get Started 2
Portal, sign-up, and the first successful call
Documentation 1
Reference material describing how the API behaves
Agent Surfaces 1
MCP servers, agent skills, and machine-readable catalogs
Design & Contract 1
Pagination, idempotency, versioning, errors, and events
Build 2
SDKs, sample code, and the tooling you integrate with
Access & Security 2
Authentication, authorization, and security posture
Operate 1
Status, limits, changes, and where to get help
Company 1
The organization behind the API