Apache Giraph website screenshot

Apache Giraph

Apache Giraph is an iterative graph processing system built for high scalability on Apache Hadoop. It is modeled after Google's Pregel and provides a simple yet flexible Java API for graph algorithms at massive scale using the Bulk Synchronous Parallel (BSP) model. Note - Apache Giraph has been retired as of 2024.

Apache Giraph publishes 2 APIs on the APIs.io network: Cluster API and Job Management API. Tagged areas include Apache, Big Data, BSP, Graph Processing, and Hadoop.

The Apache Giraph catalog on APIs.io includes 1 JSON-LD context and 2 Spectral governance rulesets.

Apache Giraph’s developer surface includes documentation, getting-started guide, and 7 more developer resources.

46.6/100 developing ▼ -5.8 Agent 25/100 agent aware Full breakdown ↓
scored 2026-07-28 · rubric v0.6
AccessFreemium
3 APIs 8 Features 5 Use Cases
ApacheBig DataBSPGraph ProcessingHadoopOpen SourceRetired

Kin Score

Kin Score Kin Score How this is scored →
scored 2026-07-28 · rubric v0.6
Composite quality — 46.6/100 · developing
Contract Quality 15.3 / 25
Developer Ergonomics 3.9 / 20
Commercial Clarity 7.9 / 20
Operational Transparency 4.8 / 13
Governance 8.3 / 12
Discoverability 6.5 / 10
Agent readiness — 25/100 · agent aware
Machine-Readable Contract 18 / 18
Agentic Access Contract 10 / 10
MCP Server 0 / 12
Machine-Readable Auth 0 / 10
Idempotency 0 / 9
Stable Error Semantics 0 / 8
Request/Response Examples 7 / 7
Rate-Limit Signaling 7 / 7
Typed Event Surface 0 / 6
Agent Skills 0 / 5
Well-Known Catalog 0 / 4
Consent & Bot Identity 0 / 3
A2A Agent Card 0 / 8
Dry-Run / Simulate Mode 0 / 4
Improve this rating by publishing the missing artifacts — every area above can be raised, and the full rubric is at apis.io/rating/. This rating is computed from github.com/api-evangelist/apache-giraph: open an issue to ask a question, or submit a pull request to add artifacts. Want it done for you? Prioritized profiling — $2,500 →

APIs 3

Individual APIs this provider publishes, each with its own machine-readable definition.

Apache Giraph Java API

Java API based on the Bulk Synchronous Parallel (BSP) model for implementing graph algorithms, with Vertex, Edge, and Master compute APIs for distributed graph processing on Had...

Apache Giraph Cluster API

The Cluster API from Apache Giraph — 1 operation(s) for cluster.

Apache Giraph Job Management API

The Job Management API from Apache Giraph — 2 operation(s) for job management.

Open Collections 1

Open, tool-agnostic API collections (OpenAPI-derived and Bruno).

Pricing Plans 1

Published pricing tiers and plan structures.

Rate Limits 1

Documented rate limits and quota policies.

Apache Giraph Rate Limits

5 limits

RATE LIMITS

FinOps 1

Cost, billing, and metering signals for API financial operations.

Features 8

Notable capabilities this provider offers.

Bulk Synchronous Parallel (BSP) Model

Google Pregel-inspired BSP computation model where vertices communicate through supersteps.

Vertex-Centric Programming

Write graph algorithms by defining per-vertex compute functions that exchange messages with neighbors.

Master Compute API

Global coordination API for aggregating results and controlling algorithm termination across supersteps.

Aggregators

Sharded aggregators for collecting global statistics across all vertices during computation.

Edge-Oriented Input

Flexible input formats for loading graphs from HDFS, Hive, Gora, and Rexster sources.

Out-of-Core Computation

Spill graph data to disk for processing graphs larger than available memory.

Hadoop Integration

Runs as a MapReduce job on Apache Hadoop YARN for resource management and fault tolerance.

Fault Tolerance

Checkpoint-based recovery for fault tolerance across superstep boundaries.

Scroll for all 8

Semantic Vocabularies 1

JSON-LD contexts and semantic vocabularies used across these APIs.

Apache Giraph Job Context

5 classes · 14 properties

JSON-LD

Spectral Rules 2

Spectral governance rulesets for linting and validating these APIs.

Apache Giraph API Rules

5 rules · 3 warnings 2 info

SPECTRAL

Apache Giraph API Rules

11 rules · 4 errors 6 warnings 1 info

SPECTRAL

JSON Schema 4

Standalone JSON Schema definitions for this provider's data models.

ApplicationInfo

12 properties

JSON SCHEMA

ApplicationResponse

1 properties

JSON SCHEMA

ApplicationsResponse

1 properties

JSON SCHEMA

ClusterMetricsResponse

1 properties

JSON SCHEMA

JSON Structure 4

JSON Structure definitions describing this provider's data shapes.

Giraph Job Application Info Structure

12 properties

JSON STRUCTURE

Giraph Job Application Response Structure

1 properties

JSON STRUCTURE

Giraph Job Applications Response Structure

1 properties

JSON STRUCTURE

Examples 4

Example request and response payloads for these APIs.

Security Posture 2

Authentication, domain security, vulnerability disclosure, and trust-center signals.

Apache Giraph Domain Security

TLSv1.3 · HSTS · DMARC

SECURITY

Apache Giraph Vulnerability Disclosure

security.txt · contact published

SECURITY

Agentic Access 1

Recommended x-agentic-access execution contracts for AI agents.

Apache Giraph Agentic Access

3 operations

3 operations · 0 acting

AGENTIC

Use Cases 5

What developers build with this provider.

Social Graph Analysis

Analyze social network connections, communities, and influence at billions-of-vertices scale (as used at Facebook).

PageRank Computation

Compute web page or entity rankings using iterative link analysis algorithms.

Shortest Path Computation

Find shortest paths between vertices for network routing and recommendation problems.

Connected Components

Identify clusters and connected components in large graphs for community detection.

Graph Machine Learning Features

Generate graph-structural features for machine learning models at scale.

Integrations 5

Pre-built integrations with other platforms and tools.

Apache Hadoop

Runs on Hadoop YARN as a MapReduce application for cluster resource management.

Apache Hive

Hive I/O module for loading graph data from Hive tables.

Apache Gora

Gora I/O module for loading graph data from various NoSQL data stores.

Rexster

Rexster graph server I/O module for loading data from TinkerPop graph databases.

Apache HBase

HBase integration for storing and loading vertex and edge data.

Resources

Get Started 1

Portal, sign-up, and the first successful call

Documentation 1

Reference material describing how the API behaves

Agent Surfaces 1

MCP servers, agent skills, and machine-readable catalogs

Design & Contract 2

Pagination, idempotency, versioning, errors, and events

Build 2

SDKs, sample code, and the tooling you integrate with

Access & Security 2

Authentication, authorization, and security posture

Source (apis.yml)

apis.yml Raw ↑
aid: apache-giraph
name: Apache Giraph
description: Apache Giraph is an iterative graph processing system built for high scalability on Apache Hadoop. It is modeled
  after Google's Pregel and provides a simple yet flexible Java API for graph algorithms at massive scale using the Bulk Synchronous
  Parallel (BSP) model. Note - Apache Giraph has been retired as of 2024.
type: Index
accessModel:
  pricing: freemium
  onboarding: unknown
  trial: false
  try_now: false
  public: false
  label: Freemium
  confidence: medium
  source:
  - plans
  generated: '2026-07-22'
  method: derived
position: Consumer
access: 3rd-Party
image: https://kinlane-images.s3.amazonaws.com/shared/apis-json/icons/apache-giraph.png
tags:
- Apache
- Big Data
- BSP
- Graph Processing
- Hadoop
- Open Source
- Retired
created: '2026-03-16'
modified: '2026-05-19'
url: https://raw.githubusercontent.com/api-evangelist/apache-giraph/refs/heads/main/apis.yml
specificationVersion: '0.19'
apis:
- aid: apache-giraph:apache-giraph-java-api
  name: Apache Giraph Java API
  description: Java API based on the Bulk Synchronous Parallel (BSP) model for implementing graph algorithms, with Vertex,
    Edge, and Master compute APIs for distributed graph processing on Hadoop.
  humanURL: https://giraph.apache.org/apidocs/
  tags:
  - BSP
  - Graph
  - Java
  - SDK
  properties:
  - type: Documentation
    url: https://giraph.apache.org/apidocs/
  - type: SDKs
    url: https://search.maven.org/artifact/org.apache.giraph/giraph-core
    title: Java SDK (Maven Central)
- aid: apache-giraph:apache-giraph-cluster-api
  name: Apache Giraph Cluster API
  description: The Cluster API from Apache Giraph — 1 operation(s) for cluster.
  humanURL: https://giraph.apache.org/quick_start.html
  baseURL: http://localhost:8088
  tags:
  - Cluster
  properties:
  - type: OpenAPI
    url: openapi/apache-giraph-cluster-api-openapi.yml
  - type: Documentation
    url: https://giraph.apache.org/quick_start.html
  - type: JSONSchema
    url: json-schema/giraph-job-application-info-schema.json
  - type: JSONLD
    url: json-ld/apache-giraph-job-context.jsonld
- aid: apache-giraph:apache-giraph-job-management-api
  name: Apache Giraph Job Management API
  description: The Job Management API from Apache Giraph — 2 operation(s) for job management.
  humanURL: https://giraph.apache.org/quick_start.html
  baseURL: http://localhost:8088
  tags:
  - Job Management
  properties:
  - type: OpenAPI
    url: openapi/apache-giraph-job-management-api-openapi.yml
  - type: Documentation
    url: https://giraph.apache.org/quick_start.html
  - type: JSONSchema
    url: json-schema/giraph-job-application-info-schema.json
  - type: JSONLD
    url: json-ld/apache-giraph-job-context.jsonld
common:
- type: AgenticAccess
  url: agentic-access/apache-giraph-agentic-access.yml
- type: VulnerabilityDisclosure
  url: security/apache-giraph-vulnerability-disclosure.yml
- type: DomainSecurity
  url: security/apache-giraph-domain-security.yml
- type: Documentation
  url: https://giraph.apache.org/
- type: GettingStarted
  url: https://giraph.apache.org/quick_start.html
- type: GitHubOrganization
  url: https://github.com/apache
- type: GitHubRepository
  url: https://github.com/apache/giraph
- type: SpectralRules
  url: rules/apache-giraph-spectral-rules.yml
- type: Vocabulary
  url: vocabulary/apache-giraph-vocabulary.yaml
- type: Features
  data:
  - name: Bulk Synchronous Parallel (BSP) Model
    description: Google Pregel-inspired BSP computation model where vertices communicate through supersteps.
  - name: Vertex-Centric Programming
    description: Write graph algorithms by defining per-vertex compute functions that exchange messages with neighbors.
  - name: Master Compute API
    description: Global coordination API for aggregating results and controlling algorithm termination across supersteps.
  - name: Aggregators
    description: Sharded aggregators for collecting global statistics across all vertices during computation.
  - name: Edge-Oriented Input
    description: Flexible input formats for loading graphs from HDFS, Hive, Gora, and Rexster sources.
  - name: Out-of-Core Computation
    description: Spill graph data to disk for processing graphs larger than available memory.
  - name: Hadoop Integration
    description: Runs as a MapReduce job on Apache Hadoop YARN for resource management and fault tolerance.
  - name: Fault Tolerance
    description: Checkpoint-based recovery for fault tolerance across superstep boundaries.
- type: UseCases
  data:
  - name: Social Graph Analysis
    description: Analyze social network connections, communities, and influence at billions-of-vertices scale (as used at
      Facebook).
  - name: PageRank Computation
    description: Compute web page or entity rankings using iterative link analysis algorithms.
  - name: Shortest Path Computation
    description: Find shortest paths between vertices for network routing and recommendation problems.
  - name: Connected Components
    description: Identify clusters and connected components in large graphs for community detection.
  - name: Graph Machine Learning Features
    description: Generate graph-structural features for machine learning models at scale.
- type: Integrations
  data:
  - name: Apache Hadoop
    description: Runs on Hadoop YARN as a MapReduce application for cluster resource management.
  - name: Apache Hive
    description: Hive I/O module for loading graph data from Hive tables.
  - name: Apache Gora
    description: Gora I/O module for loading graph data from various NoSQL data stores.
  - name: Rexster
    description: Rexster graph server I/O module for loading data from TinkerPop graph databases.
  - name: Apache HBase
    description: HBase integration for storing and loading vertex and edge data.
maintainers:
- FN: Kin Lane
  email: info@apievangelist.com