Use Case
ETL & Data Pipelines
Extracting, transforming, and loading data at scale.
47 providers
293 APIs
29 declared variants
Building and orchestrating batch and streaming pipelines that move data from operational sources into analytical stores, including scheduling, retries, and pipeline observability.
47 API providers on the APIs.io network offer etl & data pipelines. The highest-rated are Amazon EC2 Auto Scaling, Temporal, Claude, Autodesk, Amazon Managed Service for Apache Flink.
Providers
Ranked by API Evangelist rating — Exemplar and Strong are expanded by default.
Exemplar 2 Complete, well-documented, and agent-ready
Strong 14 Solid coverage with minor gaps
Claude
Anthropic's Claude AI assistant API for natural language processing and conversation.
Autodesk
Autodesk is a global leader in design, engineering, and entertainment software, providing cloud-connected p...
Amazon Managed Service for Apache Flink
Amazon Managed Service for Apache Flink is the easiest way to transform and analyze streaming data in real ...
Amazon Data Exchange
AWS Data Exchange makes it easy to find, subscribe to, and use third-party data in the cloud. Qualified dat...
Microsoft Azure Functions
Azure Functions is a serverless compute platform from Microsoft Azure enabling event-driven code execution ...
Factset
FactSet creates flexible, open data and software solutions for tens of thousands of investment professional...
Amazon Kinesis Data Firehose
Amazon Kinesis Data Firehose is the easiest way to reliably load streaming data into data lakes, data store...
Amazon Glue DataBrew
AWS Glue DataBrew is a visual data preparation tool that makes it easy for data analysts and data scientist...
IBM WebSphere
IBM WebSphere is a family of enterprise software products that provide middleware and application server ca...
123FormBuilder
123FormBuilder is an online form, survey, and workflow builder used to collect, route, and integrate submis...
Amazon Step Functions
Amazon Step Functions is a serverless workflow orchestration service that lets you coordinate distributed a...
Amazon Redshift
Amazon Redshift is a fast, fully managed cloud data warehouse that makes it simple and cost-effective to an...
Apache Airflow
Apache Airflow is an open-source platform to programmatically author, schedule, and monitor workflows, deve...
Azure Databricks
Azure Databricks is an Apache Spark-based analytics platform optimized for Microsoft Azure. It provides a c...
Developing 18 Usable, with meaningful gaps to close
Apache Oozie
Apache Oozie is a workflow scheduler system for managing Apache Hadoop jobs. It enables users to define wor...
Backpack
Backpack is a Solana-first crypto company founded by Armani Ferrante and Tristan Yver — the same team behin...
Amazon ECS
Amazon Elastic Container Service (ECS) is a fully managed container orchestration service that makes it eas...
AWS Step Functions
AWS Step Functions is a serverless orchestration service that lets you coordinate distributed applications ...
Apache Kafka
Apache Kafka is an open-source distributed event streaming platform used by thousands of companies for high...
Argo Workflows
Argo Workflows is an open-source, container-native workflow engine for orchestrating parallel jobs on Kuber...
Apache Iceberg
Apache Iceberg is an open table format for large analytic datasets that provides ACID transactions, schema ...
Apache Airflow
Apache Airflow is an open-source platform to programmatically author, schedule, and monitor workflows. Airf...
Vert.x
Eclipse Vert.x is a toolkit for building reactive applications on the JVM, providing support for multiple l...
Apache Pig
Apache Pig is a platform for analyzing large data sets that provides a high-level language (Pig Latin) for ...
Apache Flink
Apache Flink is a framework and distributed processing engine for stateful computations over unbounded and ...
Apache DolphinScheduler
Apache DolphinScheduler is a modern distributed and extensible data orchestration platform governed by the ...
Azure Logic Apps
Azure Logic Apps is a cloud platform for creating and running automated workflows that integrate apps, data...
Apache Hive
Apache Hive is a data warehouse software that facilitates reading, writing, and managing large datasets res...
DreamFactory
Automate the building, securing, and documenting of REST APIs for data products with built-in enterprise se...
Tableau Desktop
APIs and integration points for Tableau Desktop, a data visualization and business intelligence platform fr...
Apache Pulsar
Apache Pulsar is a cloud-native, distributed messaging and streaming platform that provides server-to-serve...
Power Query
Power Query is a data transformation and mashup engine used across Microsoft products including Excel, Powe...
Thin 10 Limited public surface area
Conductor
Conductor allows you to build a complex application using simple and granular tasks that do not need to be ...
Qlik Sense APIs
Collection of APIs for Qlik Sense, a business intelligence and visual analytics platform. Qlik provides RES...
Amazon Athena
Amazon Athena is an interactive query service that makes it easy to analyze data in Amazon S3 using standar...
Apache Storm
Apache Storm is a free and open-source distributed real-time computation system that makes it easy to relia...
Apache NiFi
Apache NiFi is a dataflow management system designed to automate the flow of data between systems. It provi...
Apache Mesos
Apache Mesos is a retired cluster manager (now in the Apache Attic) that provided efficient resource isolat...
Apache Arrow
Apache Arrow is a cross-language development platform for in-memory analytics developed by the Apache Softw...
Apache Camel
Apache Camel is an open-source integration framework developed by the Apache Software Foundation that imple...
Apache Beam
Apache Beam is a unified, open-source programming model developed by the Apache Software Foundation for def...
Drip
Drip is a Minneapolis-based ecommerce marketing automation platform that combines email, SMS, popups, and w...
Emerging 2 Early or largely undocumented
Minimal 1 Almost no public developer surface
What Providers Actually Declared
This use case is a canonical term. These are the free-text strings
providers wrote in their own apis.yml that map onto it.
Automated Data PipelinesBatch ProcessingBatch Processing and Job ManagementBig Data ProcessingBuild serverless data pipelines and reporting solutionsClinical Data IngestionComplex ETL PipelinesData IngestionData Ingestion PipelinesData PipelineData Pipeline AccelerationData Pipeline AutomationData Pipeline OrchestrationData Pipeline ProcessingData Processing PipelinesData TransformationData engineers building automated financial data pipelines and reportsDesign Batch ProcessingETL Pipeline AutomationETL Pipeline OrchestrationETL Pipeline ProcessingETL PipelinesETL ProcessingIoT Data IngestionReal-Time Data PipelinesReal-time market data ingestionRunning ETL pipelines for data transformationSubscriber and Event Data PipelinesSurvey Data Pipelines
Scroll for all 29