Scraping
Variants seen in the corpus:
Scrapingscraping
Providers using this tag (22)
Ranked by API Evangelist rating — Exemplar and Strong are expanded by default.
Exemplar 2 Complete, well-documented, and agent-ready
Developing 5 Usable, with meaningful gaps to close
Emerging 5 Early or largely undocumented
APIs with this tag (30)
Ranked by the provider's API Evangelist rating — the Kin Score is scored per provider, not per API, so every API of a provider shares its band. How the rating works →
Exemplar 7 Complete, well-documented, and agent-ready
Diffbot Bulk Extract APIDiffbot Bulk Extract API is a tool that allows users to extract data at scale from a variety of sources, in...Diffbot Crawl APIDiffbot Crawl API is a powerful tool that automates the process of extracting content and data from website...Diffbot Crawl/Bulk Job APIThe Diffbot Crawl/Bulk Job API is a powerful tool that allows users to automatically extract and organize l...Diffbot DQL APIThe Diffbot DQL API is a powerful tool that allows users to query and retrieve data from the web in a struc...Diffbot Enhance APIDiffbot Enhance API enhances data by providing additional context and insights. By analyzing text and image...Diffbot Extract APIDiffbot Extract API is a powerful tool that allows users to automatically extract multiple types of data fr...Diffbot Natural Language APIDiffbot Natural Language API allows users to extract and analyze textual content from websites. By utilizin...
Strong 2 Solid coverage with minor gaps
Developing 8 Usable, with meaningful gaps to close
ScrapingAntScrapingAnt is a web scraping API service that handles proxy rotation, headless browsers, and CAPTCHA solvi...ScrapingAnt MCP ServerFirst-party hosted remote MCP server at https://api.scrapingant.com/mcp/ exposing get_web_page_html, get_we...ScrapingAnt Scraping APIThe ScrapingAnt web scraping endpoint. GET/POST/PUT/PATCH/DELETE /v2/general renders a target URL in headle...AnyAPI Gateway APIUnified REST gateway to 363 normalized scraping and data APIs, published as an OpenAPI 3.1.0 document with ...Zoca Scraping APIThe Scraping API from Zoca — 6 operation(s) for scraping.Spider Cloud MCP ServerSpider's hosted Model Context Protocol server exposes 22 tools — eight core operations (crawl, scrape, sear...Spider Scraping APIExtract content from individual web pages.Firecrawl Scraping APIThe Scraping API from Firecrawl — 6 operation(s) for scraping.
Thin 7 Limited public surface area
Notte Scraping APIOne-shot scraping of a URL or raw HTML, plus AI web search.Prometheus Exposition Format / OpenMetricsThe text-based exposition format that every instrumented target exposes (typically on /metrics) and that th...Scrapfly Scraping APIThe Scraping API from Scrapfly — 1 operation(s) for scraping.DataDome Bot ProtectBot Protect is DataDome's core bot management product, scoring every request against the platform's threat ...ScraperAPIScraperAPI is a web scraping API that manages proxies, browsers, and CAPTCHAs to extract HTML from any web ...Crawlee JavaScript SDKThe Crawlee JavaScript SDK is a Node.js/TypeScript library for building reliable web scrapers and crawlers....Crawlee Python SDKThe Crawlee Python SDK is a Python library for building reliable web scrapers and crawlers. It offers Basic...
Emerging 6 Early or largely undocumented
CheerioCheerio implements a subset of core jQuery designed for the server. It parses markup into a traversable, ma...HUMAN Bot DefenderBot Defender (formerly PerimeterX Bot Defender) is HUMAN's flagship product for stopping automated traffic ...HUMAN Scraper ProtectionScraper Protection targets large-scale content scraping by automated agents, including LLM training scraper...Beautiful SoupBeautiful Soup 4 is a Python library providing a parse tree API for HTML and XML documents. It exposes Tag,...ScrapyScrapy is an open-source Python web crawling framework for extracting structured data from websites using s...PuppeteerPuppeteer is a Node.js library providing a high-level API to control headless Chrome or Chromium browsers f...
Companies reaching this through an API (6)
These companies publish an API, specification or operation carrying “Scraping” but do not classify their business under it. Listed unranked and kept out of the count above, because one tagged operation is not a statement about what a company does.
Score breakdown
Frequency
48.5
log-scaled weighted occurrences
Breadth
0.6
spread across providers
Quality lift
45.3
mean composite of providers using it
Cohesion
1.5
strength of nearest seed neighbor
Related tags
Crawling 5 co-occurrences
Proxies 6 co-occurrences
Data Extraction 13 co-occurrences
Browser Automation 7 co-occurrences
Web Scraping 5 co-occurrences
LLM 5 co-occurrences
Billing 5 co-occurrences
Data 5 co-occurrences
Where this tag comes from
Provider tag22
Api tag30
Openapi tag6
Openapi op tag25
Work with this as data
Every tag here is available over the APIs.io API and to AI agents over MCP.