Fish Audio website screenshot

Fish Audio

Fish Audio is an AI voice platform offering text-to-speech, voice cloning, speech-to-text, voice changing, and audio storytelling capabilities. The platform hosts a library of over two million voices across 30+ languages and is built around the Fish Speech open-source TTS model and the proprietary Fish Audio S2-Pro model. Fish Audio exposes a public REST API at api.fish.audio with first-party Python, Go, and TypeScript SDKs and supports voice cloning from as little as fifteen seconds of reference audio. The developer surface emphasizes ultra-low latency streaming, emotion control, and pay-as-you-go pricing for both prototype and production workloads.

Fish Audio publishes 4 APIs on the APIs.io network, including Asr API, Model API, Tts API, and 1 more. Tagged areas include Voice, Text to Speech, Speech to Text, Voice Cloning, and Audio.

Fish Audio’s developer surface includes authentication, documentation, pricing, and 10 more developer resources.

35.7/100 thin ▼ -5.4 Agent 31/100 agent aware Full breakdown ↓
scored 2026-07-28 · rubric v0.6
AccessFreeSelf serve⚡ Free to try
5 APIs
VoiceText to SpeechSpeech to TextVoice CloningAudioGenerative AIMultilingualStreamingSDKOpen Source

Kin Score

Kin Score Kin Score How this is scored →
scored 2026-07-28 · rubric v0.6
Composite quality — 35.7/100 · thin
Contract Quality 13.4 / 25
Developer Ergonomics 7.0 / 20
Commercial Clarity 7.9 / 20
Operational Transparency 3.4 / 13
Governance 0.0 / 12
Discoverability 7.4 / 10
Agent readiness — 31/100 · agent aware
Machine-Readable Contract 18 / 18
Agentic Access Contract 10 / 10
MCP Server 0 / 12
Machine-Readable Auth 10 / 10
Idempotency 0 / 9
Stable Error Semantics 0 / 8
Request/Response Examples 0 / 7
Rate-Limit Signaling 7 / 7
Typed Event Surface 0 / 6
Agent Skills 0 / 5
Well-Known Catalog 0 / 4
Consent & Bot Identity 0 / 3
A2A Agent Card 0 / 8
Dry-Run / Simulate Mode 0 / 4
Improve this rating by publishing the missing artifacts — every area above can be raised, and the full rubric is at apis.io/rating/. This rating is computed from github.com/api-evangelist/fish-audio: open an issue to ask a question, or submit a pull request to add artifacts. Want it done for you? Prioritized profiling — $2,500 →

APIs 5

Individual APIs this provider publishes, each with its own machine-readable definition.

Fish Audio API

The Fish Audio API provides RESTful access to text-to-speech, speech-to-text, voice cloning, and voice management capabilities backed by the Fish Audio S2-Pro model. Endpoints s...

Fish Audio Asr API

The Asr API from Fish Audio — 1 operation(s) for asr.

Fish Audio Model API

The Model API from Fish Audio — 2 operation(s) for model.

Fish Audio Tts API

The Tts API from Fish Audio — 2 operation(s) for tts.

Fish Audio Wallet API

The Wallet API from Fish Audio — 2 operation(s) for wallet.

Open Collections 1

Open, tool-agnostic API collections (OpenAPI-derived and Bruno).

Fish Audio API

OPEN COLLECTION

Pricing Plans 1

Published pricing tiers and plan structures.

Rate Limits 1

Documented rate limits and quota policies.

Fish Audio Rate Limits

2 limits

RATE LIMITS

FinOps 1

Cost, billing, and metering signals for API financial operations.

Security Posture 2

Authentication, domain security, vulnerability disclosure, and trust-center signals.

Fish Audio Authentication

http · 1 scheme

SECURITY

Fish Audio Domain Security

TLSv1.3 · HSTS · DMARC

SECURITY

Agentic Access 1

Recommended x-agentic-access execution contracts for AI agents.

Fish Audio Agentic Access

10 operations · 6 acting

10 operations · 6 acting

AGENTIC

Resources

Get Started 1

Portal, sign-up, and the first successful call

Documentation 1

Reference material describing how the API behaves

Agent Surfaces 2

MCP servers, agent skills, and machine-readable catalogs

Build 1

SDKs, sample code, and the tooling you integrate with

Access & Security 2

Authentication, authorization, and security posture

Operate 1

Status, limits, changes, and where to get help

Commercial 1

Pricing, plans, and the legal terms of use

Company 2

The organization behind the API

Other 2

Properties that don't map to a standard resource type

Source (apis.yml)

apis.yml Raw ↑
aid: fish-audio
name: Fish Audio
description: Fish Audio is an AI voice platform offering text-to-speech, voice cloning, speech-to-text, voice changing, and
  audio storytelling capabilities. The platform hosts a library of over two million voices across 30+ languages and is built
  around the Fish Speech open-source TTS model and the proprietary Fish Audio S2-Pro model. Fish Audio exposes a public REST
  API at api.fish.audio with first-party Python, Go, and TypeScript SDKs and supports voice cloning from as little as fifteen
  seconds of reference audio. The developer surface emphasizes ultra-low latency streaming, emotion control, and pay-as-you-go
  pricing for both prototype and production workloads.
type: Index
accessModel:
  pricing: free
  onboarding: self-serve
  trial: false
  try_now: true
  public: false
  label: Free · Self-serve signup
  confidence: high
  source:
  - plans
  - authentication
  generated: '2026-07-22'
  method: derived
position: Provider
access: 3rd-Party
image: https://kinlane-images.s3.amazonaws.com/shared/apis-json/icons/fish-audio.png
tags:
- Voice
- Text to Speech
- Speech to Text
- Voice Cloning
- Audio
- Generative AI
- Multilingual
- Streaming
- SDK
- Open Source
url: https://raw.githubusercontent.com/api-evangelist/fish-audio/refs/heads/main/apis.yml
created: '2026-05-23'
modified: '2026-05-23'
specificationVersion: '0.20'
apis:
- aid: fish-audio:fish-audio-api
  name: Fish Audio API
  description: The Fish Audio API provides RESTful access to text-to-speech, speech-to-text, voice cloning, and voice management
    capabilities backed by the Fish Audio S2-Pro model. Endpoints support streaming low-latency generation, multilingual synthesis
    across 30+ languages, emotion control, and on-the-fly custom voice creation from short reference clips. The API is consumed
    through the Fish Audio Python, Go, and TypeScript SDKs and a community of integrations including n8n.
  humanURL: https://docs.fish.audio
  baseURL: https://api.fish.audio
  tags:
  - Text to Speech
  - Voice Cloning
  - Speech to Text
  - Streaming
  - REST
  - Audio
  properties:
  - type: Documentation
    url: https://docs.fish.audio
  - type: GettingStarted
    url: https://docs.fish.audio/quickstart
  - type: Playground
    url: https://fish.audio/discovery
  - type: SDKs
    url: https://github.com/fishaudio/fish-audio-python
  - type: SDKs
    url: https://github.com/fishaudio/fish-audio-go
  - type: GitHubOrganization
    url: https://github.com/fishaudio
  features:
  - name: Text-to-Speech Generation
    description: Synthesize natural, emotionally expressive speech from text using the Fish Audio S2-Pro model across 30+
      languages.
  - name: Voice Cloning
    description: Create custom voice models from as little as 15 seconds of reference audio for downstream TTS.
  - name: Speech-to-Text Transcription
    description: Transcribe audio with multispeaker detection and emotion tagging metadata.
  - name: Streaming Audio
    description: Low-latency streaming responses suitable for real-time agent, IVR, and live narration use cases.
  - name: Emotion and Prosody Control
    description: Inline emotion tags (angry, sad, excited) and special effects (laughing, sobbing) for expressive output.
  - name: Multilingual Synthesis
    description: Native support for English, Mandarin, Japanese, Korean, and more than 25 additional languages.
  - name: Voice Library
    description: Access to a hosted library of more than two million pre-built voices for instant TTS generation.
  useCases:
  - name: Audiobook and Podcast Production
    description: Generate full-length narrated content with multi-character voices via Story Studio workflows.
  - name: Conversational Agents and IVR
    description: Power voice-first agents and interactive voice response systems with low-latency synthesis.
  - name: Gaming NPC Dialogue
    description: Create dynamic in-game character voices and barks without manual voice-over sessions.
  - name: Video and Content Localization
    description: Dub and localize video, social, and marketing content across dozens of languages.
  - name: Accessibility Tooling
    description: Embed expressive screen reading and assistive voice output in accessibility products.
  integrations:
  - name: Python SDK
  - name: Go SDK
  - name: TypeScript SDK
  - name: n8n
  - name: LangChain
  - name: Hugging Face
  - name: Discord
  authentication:
  - type: API Key
    description: Requests authenticate using a Bearer API key issued from the Fish Audio dashboard.
- aid: fish-audio:fish-audio-asr-api
  name: Fish Audio Asr API
  description: The Asr API from Fish Audio — 1 operation(s) for asr.
  humanURL: https://docs.fish.audio
  baseURL: https://api.fish.audio
  tags:
  - Asr
  properties:
  - type: OpenAPI
    url: openapi/fish-audio-asr-api-openapi.yml
- aid: fish-audio:fish-audio-model-api
  name: Fish Audio Model API
  description: The Model API from Fish Audio — 2 operation(s) for model.
  humanURL: https://docs.fish.audio
  baseURL: https://api.fish.audio
  tags:
  - Model
  properties:
  - type: OpenAPI
    url: openapi/fish-audio-model-api-openapi.yml
- aid: fish-audio:fish-audio-tts-api
  name: Fish Audio Tts API
  description: The Tts API from Fish Audio — 2 operation(s) for tts.
  humanURL: https://docs.fish.audio
  baseURL: https://api.fish.audio
  tags:
  - Tts
  properties:
  - type: OpenAPI
    url: openapi/fish-audio-tts-api-openapi.yml
- aid: fish-audio:fish-audio-wallet-api
  name: Fish Audio Wallet API
  description: The Wallet API from Fish Audio — 2 operation(s) for wallet.
  humanURL: https://docs.fish.audio
  baseURL: https://api.fish.audio
  tags:
  - Wallet
  properties:
  - type: OpenAPI
    url: openapi/fish-audio-wallet-api-openapi.yml
common:
- type: AgenticAccess
  url: agentic-access/fish-audio-agentic-access.yml
- type: DomainSecurity
  url: security/fish-audio-domain-security.yml
- type: Authentication
  url: authentication/fish-audio-authentication.yml
- type: Website
  url: https://fish.audio
- type: Documentation
  url: https://docs.fish.audio
- type: DeveloperPortal
  url: https://fish.audio/go-api
- type: Playground
  url: https://fish.audio/discovery
- type: Pricing
  url: https://fish.audio/pricing
- type: GitHubOrganization
  url: https://github.com/fishaudio
- type: OpenSourceModel
  url: https://github.com/fishaudio/fish-speech
- type: Discord
  url: https://discord.gg/Es5qTB9BcN
- type: Twitter
  url: https://twitter.com/FishAudio
- type: LlmsText
  url: https://docs.fish.audio/llms.txt
maintainers:
- FN: Kin Lane
  email: kin@apievangelist.com