ToolPicker
Fish Audio

Fish Audio

54/100

OpenAudio/Fish TTS API: SOTA voices and cloning, per-million-bytes.

fish.audio

// scorecard · 54/100

verified 1mo ago
methodology →
Agent-readiness20/35
Agent registration (api_key)6/9
Public API5/5
OpenAPI spec0/9
MCP server9/9
llms.txt0/3
Reliability & performance0/15
Actively maintained
Production-ready0/4
Updated recently
Pricing12/13
Itemized public pricing6/6
Free tier / trial4/4
Free-tier generosity2/3
Docs & DX10/12
Docs quality6/8
Public API2/2
Quickstart / examples2/2
Security & compliance8/12
SOC 26/6
ISO 270010/2
GDPR0/2
Security / trust page2/2
Openness4/10
Open source0/6
Open / un-gated API4/4
Company maturity0/3
Company size
Years operating

// rankings · cost per standard workload

S2.1 Pro/S1 $15/1M UTF-8 bytes ≈ $15/1M English chars = $15/mo (non-Latin 3-4x)

fish.audio/developers

// overview

Fish Audio provides AI Text-to-Speech via the S2.1 Pro model with emotion control and special audio tags, voice cloning, and speech-to-text capabilities. It supports real-time generation, a library of 2,000,000+ voices, and use cases including video voiceovers, audiobook narration, character voices, and conversational chatbots. The platform offers API access with Python and TypeScript SDKs plus WebSocket streaming.

Best for: Creators, developers, and teams needing expressive TTS, voice cloning, and production voice AI

// pricing

Free tier · generosity 3/5
Includes: S2.1 Pro free for developers
Cloud API$15 per million characters
  • pay-as-you-go
  • WebSocket streaming
  • direction tags

// security & compliance

SOC 2 ISO 27001 GDPR HIPAA

// features

  • Emotion/special tags like [angry], [laughing], [pause]
  • Voice cloning from short audio
  • 2,000,000+ voice library
  • 15,000+ inline direction tags
  • Sub-second latency WebSocket/REST API
  • Python and TypeScript SDKs

sources