// scorecard · 54/100
verified 1mo agoAgent-readiness20/35
Agent registration (api_key)6/9
Public API5/5
OpenAPI spec0/9
MCP server9/9
llms.txt0/3
Reliability & performance0/15
Actively maintained—
Production-ready0/4
Updated recently—
Pricing12/13
Itemized public pricing6/6
Free tier / trial4/4
Free-tier generosity2/3
Docs & DX10/12
Docs quality6/8
Public API2/2
Quickstart / examples2/2
Security & compliance8/12
SOC 26/6
ISO 270010/2
GDPR0/2
Security / trust page2/2
Openness4/10
Open source0/6
Open / un-gated API4/4
Company maturity0/3
Company size—
Years operating—
// rankings · cost per standard workload
#4 in Text-to-Speech$15/ 1M chars
S2.1 Pro/S1 $15/1M UTF-8 bytes ≈ $15/1M English chars = $15/mo (non-Latin 3-4x)
fish.audio/developers// overview
Fish Audio provides AI Text-to-Speech via the S2.1 Pro model with emotion control and special audio tags, voice cloning, and speech-to-text capabilities. It supports real-time generation, a library of 2,000,000+ voices, and use cases including video voiceovers, audiobook narration, character voices, and conversational chatbots. The platform offers API access with Python and TypeScript SDKs plus WebSocket streaming.
Best for: Creators, developers, and teams needing expressive TTS, voice cloning, and production voice AI
// pricing
Free tier · generosity 3/5
Includes: S2.1 Pro free for developers
Cloud API$15 per million characters
- pay-as-you-go
- WebSocket streaming
- direction tags
// security & compliance
✓ SOC 2✗ ISO 27001✗ GDPR✗ HIPAA
// features
- Emotion/special tags like [angry], [laughing], [pause]
- Voice cloning from short audio
- 2,000,000+ voice library
- 15,000+ inline direction tags
- Sub-second latency WebSocket/REST API
- Python and TypeScript SDKs