Apache Doris
Apache Doris is a high-performance, real-time analytical database based on MPP (Massively Parallel Processing) architecture, governed by the Apache Software Foundation. It provides MySQL-protocol-compatible SQL queries, sub-second query latency on large-scale data, columnar storage with vectorized execution, real-time upsert via Stream Load and Routine Load APIs, and federated querying over data lakes (Hive, Iceberg, Hudi). It supports both shared-nothing and storage/compute-separated deployment modes.
Apache Doris publishes 1 API on the APIs.io network. Tagged areas include Analytics, Apache, Database, Lakehouse, and MPP.
The Apache Doris catalog on APIs.io includes 1 JSON-LD context and 1 Spectral governance ruleset.
Apache Doris’ developer surface includes developer portal, documentation, getting-started guide, engineering blog, Stack Overflow tag, and 6 more developer resources.
Kin Score
APIs 1
Individual APIs this provider publishes, each with its own machine-readable definition.
Apache Doris
Apache Doris provides a MySQL-compatible protocol for SQL queries, a REST API for cluster management and monitoring, Stream Load HTTP API for real-time bulk data ingestion, Rout...
Pricing Plans 1
Published pricing tiers and plan structures.
Rate Limits 1
Documented rate limits and quota policies.
Apache Doris Rate Limits
RATE LIMITSFinOps 1
Cost, billing, and metering signals for API financial operations.
Apache Doris Finops
FINOPSFeatures 8
Notable capabilities this provider offers.
MPP Columnar Analytics
Massively parallel processing with columnar storage and vectorized execution engine for high-concurrency sub-second analytical queries.
Stream Load API
HTTP-based bulk data ingestion API that loads CSV, JSON, and Parquet data in real time with transactional guarantees.
MySQL Protocol Compatibility
Fully MySQL-wire-protocol compatible, enabling use of standard MySQL clients, drivers, and BI tools without modification.
Federated Data Lakehouse Queries
Query external data in Hive, Iceberg, Hudi, and Delta Lake tables without data movement using Multi-Catalog.
Real-Time Upsert (Unique Key Model)
Primary key based upsert model supports real-time CDC data ingestion with micro-second latency row-level updates.
Routine Load from Kafka
Continuous data ingestion from Apache Kafka topics with automatic offset management and exactly-once semantics.
Tiered Storage
Hot/warm/cold data tiering with object storage (S3, HDFS) for cost-optimized storage at scale.
MCP Server
Model Context Protocol (MCP) server enabling AI agents to query Doris databases through natural language.
Scroll for all 8
Semantic Vocabularies 1
JSON-LD contexts and semantic vocabularies used across these APIs.
Apache Doris Context
JSON-LDSpectral Rules 1
Spectral governance rulesets for linting and validating these APIs.
Apache Doris API Rules
SPECTRALJSON Schema 3
Standalone JSON Schema definitions for this provider's data models.
JSON Structure 3
JSON Structure definitions describing this provider's data shapes.
Apache Doris Routine Load Job Structure
JSON STRUCTUREApache Doris Stream Load Response Structure
JSON STRUCTUREApache Doris Table Schema Structure
JSON STRUCTUREExamples 3
Example request and response payloads for these APIs.
Security Posture 2
Authentication, domain security, vulnerability disclosure, and trust-center signals.
Use Cases 5
What developers build with this provider.
Real-Time Dashboards and Reporting
Power business intelligence dashboards with sub-second query latency on live data updated continuously.
Log and Event Analytics
Ingest and analyze high-volume log, metric, and event data in real time using inverted indexes and full-text search.
Customer Data Platform
Consolidate customer behavioral and transactional data from multiple sources for real-time segmentation and analytics.
Data Lakehouse Analytics
Federate queries across data lake (Hive, Iceberg) and operational databases without ETL movement.
Ad-Hoc Analytics
Enable data analysts to run complex exploratory SQL queries on petabyte-scale datasets with fast response times.
Integrations 6
Pre-built integrations with other platforms and tools.
Apache Flink
Official Flink Connector for reading from and writing to Doris in real-time Flink streaming pipelines.
Apache Spark
Official Spark Connector for batch ETL and analytics workflows using Apache Spark.
Apache Kafka
Kafka Connector and Routine Load for continuous real-time data ingestion from Kafka topics.
Apache Iceberg / Hudi / Hive
Multi-Catalog feature enables federated queries over Iceberg, Hudi, and Hive Metastore data lakes.
Kubernetes
Official Kubernetes Operator for automated Doris cluster lifecycle management.
OpenTelemetry
OpenTelemetry demo integration for observability and tracing in Doris deployments.
Resources
Get Started 2
Portal, sign-up, and the first successful call
Documentation 1
Reference material describing how the API behaves
Agent Surfaces 1
MCP servers, agent skills, and machine-readable catalogs
Design & Contract 1
Pagination, idempotency, versioning, errors, and events
Build 2
SDKs, sample code, and the tooling you integrate with
Access & Security 2
Authentication, authorization, and security posture
Operate 1
Status, limits, changes, and where to get help
Company 1
The organization behind the API