Apache Hive website screenshot

Apache Hive

Apache Hive is a data warehouse software that facilitates reading, writing, and managing large datasets residing in distributed storage using SQL. It provides a SQL-like interface called HiveQL for querying data stored in Hadoop, along with a WebHCat REST API for job submission and metastore access.

Apache Hive publishes 3 APIs on the APIs.io network: Databases API, Jobs API, and Tables API. Tagged areas include Apache, Big Data, Data Warehouse, ETL, and Hadoop.

The Apache Hive catalog on APIs.io includes 1 JSON-LD context and 2 Spectral governance rulesets.

Apache Hive’s developer surface includes documentation, getting-started guide, and 8 more developer resources.

45.9/100 developing ▼ -5.9 Agent 22/100 agent aware Full breakdown ↓
scored 2026-07-28 · rubric v0.6
AccessFreemium
4 APIs 8 Features 5 Use Cases
ApacheBig DataData WarehouseETLHadoopOpen SourceSQL

Kin Score

Kin Score Kin Score How this is scored →
scored 2026-07-28 · rubric v0.6
Composite quality — 45.9/100 · developing
Contract Quality 14.6 / 25
Developer Ergonomics 3.9 / 20
Commercial Clarity 7.9 / 20
Operational Transparency 4.8 / 13
Governance 8.3 / 12
Discoverability 6.5 / 10
Agent readiness — 22/100 · agent aware
Machine-Readable Contract 18 / 18
Agentic Access Contract 10 / 10
MCP Server 0 / 12
Machine-Readable Auth 0 / 10
Idempotency 0 / 9
Stable Error Semantics 0 / 8
Request/Response Examples 0 / 7
Rate-Limit Signaling 7 / 7
Typed Event Surface 0 / 6
Agent Skills 0 / 5
Well-Known Catalog 0 / 4
Consent & Bot Identity 0 / 3
A2A Agent Card 0 / 8
Dry-Run / Simulate Mode 0 / 4
Improve this rating by publishing the missing artifacts — every area above can be raised, and the full rubric is at apis.io/rating/. This rating is computed from github.com/api-evangelist/apache-hive: open an issue to ask a question, or submit a pull request to add artifacts. Want it done for you? Prioritized profiling — $2,500 →

APIs 4

Individual APIs this provider publishes, each with its own machine-readable definition.

Apache Hive JDBC API

JDBC interface to HiveServer2 for standard SQL client connectivity, supporting parameterized queries, result sets, and connection pooling from Java and ODBC-bridge applications.

Apache Hive Databases API

Database metadata operations

Apache Hive Jobs API

Hive job submission and monitoring

Apache Hive Tables API

Table metadata operations

Open Collections 1

Open, tool-agnostic API collections (OpenAPI-derived and Bruno).

Pricing Plans 1

Published pricing tiers and plan structures.

Rate Limits 1

Documented rate limits and quota policies.

Apache Hive Rate Limits

5 limits

RATE LIMITS

FinOps 1

Cost, billing, and metering signals for API financial operations.

Features 8

Notable capabilities this provider offers.

HiveQL SQL Interface

SQL-like query language for reading, writing, and aggregating data stored in distributed storage.

WebHCat REST API

HTTP REST API (Templeton) for DDL operations, job submission, and metastore metadata access.

HiveServer2 JDBC/ODBC

Thrift-based server with JDBC and ODBC drivers for standard SQL client connectivity.

Hive Metastore

Central repository for table schema, partition metadata, and storage location information.

Partitioning

Partition tables by column values for efficient query pruning and data organization.

ORC and Parquet Storage

Optimized columnar storage formats with predicate pushdown and compression support.

ACID Transactions

Full ACID transaction support for inserts, updates, and deletes on managed ORC tables.

Vectorized Query Execution

Batch processing of rows in CPU register-width vectors for improved query throughput.

Scroll for all 8

Semantic Vocabularies 1

JSON-LD contexts and semantic vocabularies used across these APIs.

Apache Hive Webhcat Context

21 classes · 0 properties

JSON-LD

Spectral Rules 2

Spectral governance rulesets for linting and validating these APIs.

Apache Hive API Rules

5 rules · 3 warnings 2 info

SPECTRAL

Apache Hive API Rules

13 rules · 2 errors 9 warnings 2 info

SPECTRAL

JSON Schema 6

Standalone JSON Schema definitions for this provider's data models.

Column

3 properties

JSON SCHEMA

Database

5 properties

JSON SCHEMA

Job

6 properties

JSON SCHEMA

Partition

5 properties

JSON SCHEMA

QueryResult

4 properties

JSON SCHEMA

Table

8 properties

JSON SCHEMA

JSON Structure 6

JSON Structure definitions describing this provider's data shapes.

Hive Webhcat Column Structure

3 properties

JSON STRUCTURE

Hive Webhcat Database Structure

5 properties

JSON STRUCTURE

Hive Webhcat Job Structure

6 properties

JSON STRUCTURE

Hive Webhcat Partition Structure

5 properties

JSON STRUCTURE

Hive Webhcat Queryresult Structure

4 properties

JSON STRUCTURE

Hive Webhcat Table Structure

8 properties

JSON STRUCTURE

Examples 6

Example request and response payloads for these APIs.

Hive Webhcat Job Example

6 fields

EXAMPLE

Security Posture 2

Authentication, domain security, vulnerability disclosure, and trust-center signals.

Apache Hive Domain Security

TLSv1.3 · HSTS · DMARC

SECURITY

Apache Hive Vulnerability Disclosure

security.txt · contact published

SECURITY

Agentic Access 1

Recommended x-agentic-access execution contracts for AI agents.

Apache Hive Agentic Access

7 operations · 1 acting

7 operations · 1 acting

AGENTIC

Use Cases 5

What developers build with this provider.

Data Warehouse Analytics

Run SQL analytics on petabyte-scale datasets stored in HDFS or object storage.

ETL Pipeline Orchestration

Use HiveQL scripts to transform and load data between raw and curated data lake zones.

Ad-Hoc Data Exploration

Query structured data interactively using Beeline or JDBC-connected BI tools.

Log Analysis

Parse and aggregate application logs stored as text or JSON in HDFS using Hive SerDes.

Data Catalog Integration

Use the Hive Metastore as a shared schema registry for Spark, Flink, and Presto.

Resources

Get Started 1

Portal, sign-up, and the first successful call

Documentation 1

Reference material describing how the API behaves

Agent Surfaces 1

MCP servers, agent skills, and machine-readable catalogs

Design & Contract 2

Pagination, idempotency, versioning, errors, and events

Build 2

SDKs, sample code, and the tooling you integrate with

Access & Security 2

Authentication, authorization, and security posture

Company 1

The organization behind the API

Source (apis.yml)

apis.yml Raw ↑
aid: apache-hive
name: Apache Hive
description: Apache Hive is a data warehouse software that facilitates reading, writing, and managing large datasets residing
  in distributed storage using SQL. It provides a SQL-like interface called HiveQL for querying data stored in Hadoop, along
  with a WebHCat REST API for job submission and metastore access.
type: Index
accessModel:
  pricing: freemium
  onboarding: unknown
  trial: false
  try_now: false
  public: false
  label: Freemium
  confidence: medium
  source:
  - plans
  generated: '2026-07-22'
  method: derived
position: Consumer
access: 3rd-Party
image: https://kinlane-images.s3.amazonaws.com/shared/apis-json/icons/apache-hive.png
tags:
- Apache
- Big Data
- Data Warehouse
- ETL
- Hadoop
- Open Source
- SQL
created: '2026-03-16'
modified: '2026-05-19'
url: https://raw.githubusercontent.com/api-evangelist/apache-hive/refs/heads/main/apis.yml
specificationVersion: '0.19'
apis:
- aid: apache-hive:apache-hive-jdbc
  name: Apache Hive JDBC API
  description: JDBC interface to HiveServer2 for standard SQL client connectivity, supporting parameterized queries, result
    sets, and connection pooling from Java and ODBC-bridge applications.
  humanURL: https://cwiki.apache.org/confluence/display/Hive/HiveServer2+Clients#HiveServer2Clients-JDBC
  tags:
  - JDBC
  - SQL
  - SDK
  properties:
  - type: Documentation
    url: https://cwiki.apache.org/confluence/display/Hive/HiveServer2+Clients#HiveServer2Clients-JDBC
  - type: SDKs
    url: https://search.maven.org/artifact/org.apache.hive/hive-jdbc
    title: Java JDBC Driver (Maven Central)
- aid: apache-hive:apache-hive-databases-api
  name: Apache Hive Databases API
  description: Database metadata operations
  humanURL: https://cwiki.apache.org/confluence/display/Hive/WebHCat
  baseURL: http://localhost:50111/templeton/v1
  tags:
  - Databases
  properties:
  - type: OpenAPI
    url: openapi/apache-hive-databases-api-openapi.yml
  - type: Documentation
    url: https://cwiki.apache.org/confluence/display/Hive/WebHCat
  - type: JSONSchema
    url: json-schema/hive-webhcat-table-schema.json
  - type: JSONLD
    url: json-ld/apache-hive-webhcat-context.jsonld
- aid: apache-hive:apache-hive-jobs-api
  name: Apache Hive Jobs API
  description: Hive job submission and monitoring
  humanURL: https://cwiki.apache.org/confluence/display/Hive/WebHCat
  baseURL: http://localhost:50111/templeton/v1
  tags:
  - Jobs
  properties:
  - type: OpenAPI
    url: openapi/apache-hive-jobs-api-openapi.yml
  - type: Documentation
    url: https://cwiki.apache.org/confluence/display/Hive/WebHCat
  - type: JSONSchema
    url: json-schema/hive-webhcat-table-schema.json
  - type: JSONLD
    url: json-ld/apache-hive-webhcat-context.jsonld
- aid: apache-hive:apache-hive-tables-api
  name: Apache Hive Tables API
  description: Table metadata operations
  humanURL: https://cwiki.apache.org/confluence/display/Hive/WebHCat
  baseURL: http://localhost:50111/templeton/v1
  tags:
  - Tables
  properties:
  - type: OpenAPI
    url: openapi/apache-hive-tables-api-openapi.yml
  - type: Documentation
    url: https://cwiki.apache.org/confluence/display/Hive/WebHCat
  - type: JSONSchema
    url: json-schema/hive-webhcat-table-schema.json
  - type: JSONLD
    url: json-ld/apache-hive-webhcat-context.jsonld
common:
- type: AgenticAccess
  url: agentic-access/apache-hive-agentic-access.yml
- type: VulnerabilityDisclosure
  url: security/apache-hive-vulnerability-disclosure.yml
- type: DomainSecurity
  url: security/apache-hive-domain-security.yml
- type: LinkedIn
  url: https://www.linkedin.com/company/apache-hive
- type: Documentation
  url: https://cwiki.apache.org/confluence/display/Hive/Home
- type: GettingStarted
  url: https://cwiki.apache.org/confluence/display/Hive/GettingStarted
- type: GitHubOrganization
  url: https://github.com/apache
- type: GitHubRepository
  url: https://github.com/apache/hive
- type: SpectralRules
  url: rules/apache-hive-spectral-rules.yml
- type: Vocabulary
  url: vocabulary/apache-hive-vocabulary.yaml
- type: Features
  data:
  - name: HiveQL SQL Interface
    description: SQL-like query language for reading, writing, and aggregating data stored in distributed storage.
  - name: WebHCat REST API
    description: HTTP REST API (Templeton) for DDL operations, job submission, and metastore metadata access.
  - name: HiveServer2 JDBC/ODBC
    description: Thrift-based server with JDBC and ODBC drivers for standard SQL client connectivity.
  - name: Hive Metastore
    description: Central repository for table schema, partition metadata, and storage location information.
  - name: Partitioning
    description: Partition tables by column values for efficient query pruning and data organization.
  - name: ORC and Parquet Storage
    description: Optimized columnar storage formats with predicate pushdown and compression support.
  - name: ACID Transactions
    description: Full ACID transaction support for inserts, updates, and deletes on managed ORC tables.
  - name: Vectorized Query Execution
    description: Batch processing of rows in CPU register-width vectors for improved query throughput.
- type: UseCases
  data:
  - name: Data Warehouse Analytics
    description: Run SQL analytics on petabyte-scale datasets stored in HDFS or object storage.
  - name: ETL Pipeline Orchestration
    description: Use HiveQL scripts to transform and load data between raw and curated data lake zones.
  - name: Ad-Hoc Data Exploration
    description: Query structured data interactively using Beeline or JDBC-connected BI tools.
  - name: Log Analysis
    description: Parse and aggregate application logs stored as text or JSON in HDFS using Hive SerDes.
  - name: Data Catalog Integration
    description: Use the Hive Metastore as a shared schema registry for Spark, Flink, and Presto.
- type: Integrations
  data:
  - name: Apache Hadoop HDFS
    description: Hive reads and writes data stored in HDFS as the primary storage layer.
  - name: Apache Spark
    description: Spark uses the Hive Metastore for table discovery and supports Hive UDFs.
  - name: Apache HBase
    description: Hive HBase storage handler enables HiveQL queries against HBase tables.
  - name: Apache Tez
    description: Apache Tez DAG execution engine replaces MapReduce for faster Hive query processing.
  - name: Presto / Trino
    description: Presto and Trino use the Hive Metastore for table metadata in federated SQL queries.
- type: Integrations
  url: https://cwiki.apache.org/confluence/display/connectors
integrations:
- name: Apache Connectors Framework
- name: Please wait
- name: 'User icon: Anonymous'
- name: 'User icon: kwright@metacarta.com'
maintainers:
- FN: Kin Lane
  email: info@apievangelist.com