Flexai Plans
FlexAI is usage-priced. Token Factory (serverless) is billed per model by usage (per million tokens for text/vision, per image, per generated video, per minute of audio, per million characters for TTS). Dedicated endpoints are priced per GPU-hour metered per second. New accounts receive $10 free credit; request rate limits are tiered free (10 rpm) vs. paid (100 rpm).
Flexai Plans is the machine-readable pricing-plan profile for FlexAI on the APIs.io network, conforming to the API Commons Plans specification.
It defines 3 plans, covering usage-based and custom tiers, with named plans including Token Factory (Serverless), Dedicated Endpoints, AI Factory (Private Cloud).
Tagged areas include Inference, Usage Based, and GPU Compute.
Plans
Pay-per-use serverless inference across open models, priced per model.
Reserved GPU throughput for your models and fine-tunes, priced per GPU-hour metered per second.
FlexAI platform on your own hardware — VPC, on-prem, or air-gapped. Custom pricing; contact sales.