Groq · FinOps Profile

Groq Finops

FinOps view of GroqCloud spend. Groq bills usage-based per-token rates for chat / vision / reasoning per model, per-million-character rates for TTS, per-hour transcription rates for STT, per-call or per-hour rates for tools, and a 50% Batch discount. Prompt Caching gives 50% off cached input tokens.

Groq Finops is the FinOps profile for Groq on the APIs.io network, aligned with the FinOps Foundation Framework.

It defines 9 billable meters, billed in USD, on a monthly cycle, and pricing category usage-based.

The profile maps 8 FOCUS columns for cost-allocation reporting.

Tagged areas include AI, LLM, Inference, LPU, and Low Latency.

Category: AI and Machine Learning Pricing: Usage-Based Billing: Monthly FOCUS v1.3
AILLMInferenceLPULow LatencyFinOpsCost ManagementFOCUS

Framework Alignment

Framework
Data Spec

Charge Categories

UsagePurchaseAdjustment

FOCUS Columns

BillingCurrency
USD
ChargeCategory
Usage
InvoiceIssuerName
Groq
PricingCategory
Usage-Based
ProviderName
Groq
PublisherName
Groq
ServiceCategory
AI and Machine Learning
ServiceName
GroqCloud

Meters

input_tokens
Unit: tokens
Tokens sent in chat / vision / reasoning requests, billed per 1M tokens per model.
cached_input_tokens
Unit: tokens
Cached-input tokens billed at 50% of the standard input rate.
output_tokens
Unit: tokens
Tokens generated, billed per 1M tokens per model.
tts_characters
Unit: characters
TTS characters synthesized, billed per 1M characters per voice/model.
stt_audio_hours
Unit: hours
Audio hours transcribed, billed per hour per Whisper variant.
tool_invocations
Unit: invocations
Tool calls (web search, Wolfram) priced per 1,000 invocations.
tool_compute_hours
Unit: hours
Tool compute hours (e.g., Code Execution at $0.18/hr).
batch_tokens
Unit: tokens
Tokens consumed via the Batch API at 50% discount.
flex_tokens
Unit: tokens
Tokens consumed via Flex Processing tier at relaxed-latency discount.

Sources

Work with this as data

Every finops artifact here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for finops

4 MCP tools reach this
  • find_finopsBrowse and filter every finops artifact in the catalog.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This finops artifact
curl "https://apis.io/api/v1/finops/groq-finops"
All finops
curl "https://apis.io/api/v1/finops?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no form to fill in. Signing in shares your email address with us — we store it to create your key and to recognise you if you sign in with another provider. See our Privacy Policy and Terms.

A second provider on the same verified email joins the account you already have.