Apache Pig
Apache Pig is a platform for analyzing large data sets that provides a high-level language (Pig Latin) for expressing data analysis programs. It compiles Pig Latin programs into MapReduce/Tez jobs and runs them on Hadoop clusters.
Apache Pig publishes 2 APIs on the APIs.io network: Jobs API and Scripts API. Tagged areas include Big Data, Data Analysis, ETL, Hadoop, and Scripting.
The Apache Pig catalog on APIs.io includes 1 JSON-LD context and 2 Spectral governance rulesets.
Apache Pig’s developer surface includes documentation and 7 more developer resources.
Kin Score
APIs 2
Individual APIs this provider publishes, each with its own machine-readable definition.
Apache Pig Jobs API
The Jobs API from Apache Pig — 3 operation(s) for jobs.
Apache Pig Scripts API
The Scripts API from Apache Pig — 1 operation(s) for scripts.
Pricing Plans 1
Published pricing tiers and plan structures.
Rate Limits 1
Documented rate limits and quota policies.
Apache Pig Rate Limits
RATE LIMITSFinOps 1
Cost, billing, and metering signals for API financial operations.
Apache Pig Finops
FINOPSFeatures 6
Notable capabilities this provider offers.
Pig Latin Language
High-level dataflow language for expressing data transformations
MapReduce/Tez Backend
Compiles Pig Latin to MapReduce or Apache Tez execution plans
UDF Support
User-defined functions in Java, Python, JavaScript, and Ruby
Streaming
Process data through external programs using STREAM operator
Schema Evolution
Flexible schema handling for semi-structured data
Optimization
Automatic logical and physical plan optimization
Semantic Vocabularies 1
JSON-LD contexts and semantic vocabularies used across these APIs.
Apache Pig Context
JSON-LDSpectral Rules 2
Spectral governance rulesets for linting and validating these APIs.
Apache Pig API Rules
SPECTRALApache Pig API Rules
SPECTRALJSON Schema 7
Standalone JSON Schema definitions for this provider's data models.
JobList
JSON SCHEMAJobLogs
JSON SCHEMAJobRequest
JSON SCHEMAJob
JSON SCHEMAScriptRequest
JSON SCHEMAValidationError
JSON SCHEMAValidationResult
JSON SCHEMAScroll for all 7
JSON Structure 7
JSON Structure definitions describing this provider's data shapes.
Apache Pig Job List Structure
JSON STRUCTUREApache Pig Job Logs Structure
JSON STRUCTUREApache Pig Job Request Structure
JSON STRUCTUREApache Pig Job Structure
JSON STRUCTUREApache Pig Script Request Structure
JSON STRUCTUREApache Pig Validation Error Structure
JSON STRUCTUREApache Pig Validation Result Structure
JSON STRUCTUREScroll for all 7
Examples 7
Example request and response payloads for these APIs.
Scroll for all 7
Security Posture 2
Authentication, domain security, vulnerability disclosure, and trust-center signals.
Agentic Access 1
Recommended x-agentic-access execution contracts for AI agents.
Use Cases 4
What developers build with this provider.
ETL Pipelines
Build data transformation pipelines from raw logs to structured data
Ad-hoc Data Analysis
Analyze large datasets with ad-hoc Pig Latin queries
Data Preparation
Clean and prepare data for machine learning workflows
Log Processing
Process and aggregate web server and application logs
Integrations 5
Pre-built integrations with other platforms and tools.
Apache Hadoop
Native MapReduce execution on YARN/HDFS
Apache Tez
High-performance Tez execution engine support
Apache HBase
HBase storage handler for reading/writing HBase tables
Apache Hive
HCatalog integration for Hive metastore access
Amazon S3
S3 input/output for cloud-based data processing
Resources
Documentation 1
Reference material describing how the API behaves
Agent Surfaces 1
MCP servers, agent skills, and machine-readable catalogs
Design & Contract 3
Pagination, idempotency, versioning, errors, and events
Build 1
SDKs, sample code, and the tooling you integrate with
Access & Security 2
Authentication, authorization, and security posture