ToolPicker
Cartesia

Cartesia

46/100

Real-time low-latency TTS (Sonic) for voice agents, ~40-90ms.

cartesia.ai

// scorecard · 46/100

verified 1mo ago
methodology →
Agent-readiness20/35
Agent registration (api_key)6/9
Public API5/5
OpenAPI spec0/9
MCP server9/9
llms.txt0/3
Reliability & performance0/15
Actively maintained
Production-ready0/4
Updated recently
Pricing12/13
Itemized public pricing6/6
Free tier / trial4/4
Free-tier generosity2/3
Docs & DX2/12
Docs quality
Public API2/2
Quickstart / examples0/2
Security & compliance8/12
SOC 26/6
ISO 270010/2
GDPR0/2
Security / trust page2/2
Openness4/10
Open source0/6
Open / un-gated API4/4
Company maturity0/3
Company size
Years operating

// rankings · cost per standard workload

Sonic 1 credit/char; $50/1M (PAYG)

cartesia.ai/pricing

// overview

Cartesia offers real-time speech models Sonic-3.5 (text-to-speech) and Ink-2 (speech-to-text) purpose-built for voice agents, along with the Line platform to build and deploy enterprise voice agents. Models run on State Space Models (SSMs) for low latency and synchronous interactions, with deployment options across cloud, on-premise, and on-device. It targets use cases in finance, healthcare, government, and customer support with features like voice cloning.

Best for: Enterprise voice agents and real-time speech/transcription applications in regulated industries

// pricing

Free tier · generosity 3/5
Includes: 20K credits / month · $1 prepaid agents / month
Limits: Limited to basic features
Free$0 /mo
  • 20K credits / month
  • $1 prepaid agents / month
Pro$5 /mo
  • 100K credits / month
  • Commercial use license
Startup$49 /mo
  • 1.25M credits / month
  • Pro voice cloning
Scale$299 /mo
  • 8M credits / month
  • Priority support
  • High concurrency limits
EnterpriseCustom
  • Custom credits & agent usage
  • DPAs and BAAs
  • SSO
  • Shared Slack channel

// security & compliance

SOC 2 ISO 27001 GDPR HIPAA

// features

  • Sonic-3.5 text-to-speech
  • Ink-2 streaming speech-to-text
  • Line voice agents platform
  • Instant and professional voice cloning
  • Cloud/on-premise/on-device deployment
  • Built on State Space Models (SSMs) including Mamba

sources