emerging · 0.5

Evaluation

Score 22.4 / 100 59 providers 31 APIs Search apis.io →
Business capabilities this tag reaches
Information & Data Management 1
Variants seen in the corpus: Evaluationevaluation

Providers using this tag (59)

Ranked by API Evangelist rating — Exemplar and Strong are expanded by default.

Exemplar 1 Complete, well-documented, and agent-ready
Strong 2 Solid coverage with minor gaps
Developing 28 Usable, with meaningful gaps to close
CovalCompanyAI AgentsVoice AITesting20 APIs53.9ConvaiArtificial IntelligenceConversational AICharactersNPCs10 APIs53.7Arize AILLM ObservabilityML MonitoringOpen-SourceOpenTelemetry1 API49.5ZoomChatCollaborationCommunicationsMeetings12 APIs49.5CometCompanyAi Enterprise SoftwareLLM ObservabilityLLMOps1 API49.3ScorecardCompanyArtificial IntelligenceAgents1 API49.2Microsoft Windows 10DesktopOperating SystemUWPWin3217 APIs49.0AvitoCompanyConsumerClassifiedsMarketplace25 APIs48.4OpikLLMObservabilityTracing1 API47.9VijilCompanyArtificial IntelligenceAI AgentsAgent Security1 API46.6BluejayCompanyArtificial IntelligenceAI AgentsVoice AI1 API45.8BraintrustArtificial IntelligenceLLMObservability1 API45.6Weights and BiasesMLOpsExperiment TrackingLLM ObservabilityModel Registry1 API44.1BraintrustArtificial IntelligenceLLMObservability1 API44.0TraceloopLLM ObservabilityOpenTelemetryAI MonitoringTracing1 API43.9RapidataCompanyHuman FeedbackData LabelingAnnotation1 API43.8Parea AILLMObservabilityTesting1 API42.5EigenpalCompanyDocument ProcessingArtificial IntelligenceWorkflow-Automation1 API41.5Log10LLMLoggingObservability1 API41.4DoclingDocumentsParsingPDFOCR2 APIs41.3Bespoke LabsCompanyArtificial IntelligenceMachine-LearningLLM1 API41.1Athina AIArtificial IntelligenceLLMObservability1 API40.8SplitExperimentationFeature FlagsFeature ManagementRollouts2 APIs40.4Literal AIArtificial IntelligenceLLMObservability1 API40.1FreeplayArtificial IntelligenceLLMObservability1 API39.8OpenlayerArtificial IntelligenceTestingObservability1 API39.8PromptLayerArtificial IntelligenceLLMPrompt EngineeringPrompt Management1 API39.8KluArtificial IntelligenceLLMLLM App PlatformPrompt Engineering1 API39.4
Thin 16 Limited public surface area
Emerging 11 Early or largely undocumented
Minimal 1 Almost no public developer surface

APIs with this tag (31)

Ranked by the provider's API Evangelist rating — the Kin Score is scored per provider, not per API, so every API of a provider shares its band. How the rating works →

Exemplar 1 Complete, well-documented, and agent-ready
Strong 1 Solid coverage with minor gaps
Developing 15 Usable, with meaningful gaps to close
Convai Evaluation APIThe Evaluation API from Convai — 1 operation(s) for evaluation.53.7AlyxAlyx is Arize's AI engineering agent that helps developers debug traces, create evaluators, build dashboard...49.5Arize AXArize AX is the commercial AI engineering platform covering tracing, evaluation, experiments, prompt manage...49.5PhoenixPhoenix is Arize's open-source LLM observability platform offering local tracing, evaluation, experiments, ...49.5Zoom Quality Management APIThe Zoom Quality Management API is designed to help contact centers track and analyze customer interactions...49.5Microsoft Windows 10 Evaluation APIThe Evaluation API from Microsoft Windows 10 — 1 operation(s) for evaluation.49.0Avito Evaluation APIThe Evaluation API from Avito — 5 operation(s) for evaluation.48.4Braintrust APIThe Braintrust REST API provides programmatic access to projects, experiments, datasets, prompts, functions...45.6W&B Weave (LLM Observability)LLM observability and evaluation platform providing tracing, output evaluation, cost estimation, prompt pla...44.1Rapidata Evaluation APIThe Evaluation API from Rapidata — 1 operation(s) for evaluation.43.8Eigenpal Evaluation APIManage datasets, examples, evaluators, experiment batches, evaluator scores, and run promotion workflows.41.5Log10 Evaluation APIAPI for running automated evaluations and benchmarking logged completions across multiple LLM providers, ge...41.4Log10 Feedback APIFeedback41.4Docling EvalEnd-to-end evaluation framework for document parsing models and services. Provides standard datasets and me...41.3Split Evaluation APIEndpoints for evaluating feature flags and retrieving treatment values for given keys and feature flag names.40.4
Thin 9 Limited public surface area
Emerging 5 Early or largely undocumented

Score breakdown

Frequency
53.1
log-scaled weighted occurrences
Breadth
1.6
spread across providers
Quality lift
43.6
mean composite of providers using it
Cohesion
0.0
strength of nearest seed neighbor

Related tags

Tracing 13 co-occurrences Prompts 10 co-occurrences LLM Observability 7 co-occurrences Experiments 9 co-occurrences Prompt Management 6 co-occurrences Benchmarks 7 co-occurrences Guardrails 7 co-occurrences LLM 27 co-occurrences

Where this tag comes from

Provider tag41
Api tag31
Openapi tag9
Openapi op tag36

Cohort brief

Auto-generated

The 59 providers in the APIs.io catalog tagged Evaluation, scored on the Kin Score. Every figure below is computed from the catalog — nothing here is written.

Providers
59
all scored
Mean Kin Score
36.1
+15.4 vs catalog 20.7
Mean Agent Readiness
20.2
+10.5 vs catalog 9.7
Spread
5–64.2
median 37.9 · σ 13.2
How the 59 split by band
Exemplar 0Strong 5Developing 19Thin 20Emerging 13Minimal 2
Facet averages, against the whole catalog
FacetThis cohortCatalogDifferenceScored
Contract Quality 42.9 17.5 +25.4 59
Developer Ergonomics 36.1 17.9 +18.2 59
Operational Transparency 25.5 11.3 +14.2 59
Access Clarity 33.6 22.3 +11.3 59
Discoverability 70.9 60.4 +10.5 59
Contract Governance 8.6 6.4 +2.2 59

A facet is averaged over the members that carry it, not over the whole cohort — the “Scored” column is that count. Averaging an absent facet as zero would score our own coverage gaps as the providers’ posture.

What this cohort publishes
ArtifactThis cohortCatalogDifference
MCP server (any) 27% 16% +11
MCP server (first-party) 12% 7% +5
Agent Skills 0% 0% 0
OAuth scopes 7% 9% -2
Security 98% 88% +10
Arazzo workflows 10% 2% +8
Governance rules 25% 13% +12

mcp_pct counts any mcp/ artifact including ones API Evangelist derived from the provider OpenAPI; mcp_first_party_pct counts only servers the provider publishes. Prefer the latter.

Top by Kin Score
  1. 1 OpenAI 64.2
  2. 2 Runloop 61.6
  3. 3 Amplitude 57.1
  4. 4 Coval 55.3
  5. 5 Convai 54.3
  6. 6 Scorecard 51.2
  7. 7 Zoom 51.1
  8. 8 Avito 49.8
  9. 9 Bluejay 49.8
  10. 10 Comet 49
Top by Agent Readiness
  1. 1 Coval 43.6
  2. 2 Avito 42.6
  3. 3 Rapidata 41.7
  4. 4 Scorecard 41.2
  5. 5 Eigenpal 40.8
  6. 6 Bluejay 38.6
  7. 7 OpenAI 37.7
  8. 8 Zoom 34.1
  9. 9 Together AI 33.5
  10. 10 Amplitude 32.9
All 59 members of this roster resolve to a scored provider in the catalog.Generated from the catalog build of 25 August 2026, across 26891 providers, using the same computation served by the APIs.io cohort API. Briefs are published for rosters of 5 or more scored providers; below that a distribution is not meaningful.

Work with this as data

Every tag here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for tags

7 MCP tools reach this
  • find_tagsBrowse and filter every tag in the catalog.
  • get_cohortThis tag as a scored cohort — every provider carrying it, with scores.
  • cohort_statsPRO — the distribution across this tag: mean, median, band split, adoption rates.
  • cohort_rankingsPRO — the leaderboard, on composite AND agent-readiness axes.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This tag
curl "https://apis.io/api/v1/tags/evaluation"
All tags
curl "https://apis.io/api/v1/tags?limit=25"
As a scored cohort
curl "https://apis.io/api/v1/cohorts/tag/evaluation"
The distribution (Pro)
curl "https://apis.io/api/v1/cohorts/tag/evaluation/stats" \
  -H "X-API-Key: $APIS_IO_KEY"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no email required.

A second provider on the same verified email joins the account you already have.