ToolPicker
Deepgram

Deepgram

46/100

Nova-3 speech-to-text for pre-recorded & streaming audio; fast, low-cost.

deepgram.com

// scorecard · 46/100

verified 1mo ago
methodology →
Agent-readiness20/35
Agent registration (api_key)6/9
Public API5/5
OpenAPI spec0/9
MCP server9/9
llms.txt0/3
Reliability & performance0/15
Actively maintained
Production-ready0/4
Updated recently
Pricing12/13
Itemized public pricing6/6
Free tier / trial4/4
Free-tier generosity2/3
Docs & DX2/12
Docs quality
Public API2/2
Quickstart / examples0/2
Security & compliance8/12
SOC 26/6
ISO 270010/2
GDPR0/2
Security / trust page2/2
Openness4/10
Open source0/6
Open / un-gated API4/4
Company maturity0/3
Company size
Years operating

// rankings · cost per standard workload

#11 in Transcription$25.8/ 100h audio

Nova-3 $0.0043/min x6000 = $25.80/mo

deepgram.com/pricing

// overview

Deepgram provides real-time and batch APIs for speech-to-text (STT), text-to-speech (TTS), Voice Agent, and Audio Intelligence. It offers a unified Voice Agent API combining STT, TTS, and LLM orchestration in one call, with models like Flux (conversational, multilingual, turn detection) and Nova (high-accuracy transcription). Available in cloud and self-hosted, with support for custom models and enterprise deployments.

Best for: Developers, product teams, platforms, and enterprises building scalable voice AI applications

// pricing

Free tier · generosity 3/5
Includes: Pay As You Go access to public models
Limits: STT up to 50 concurrent REST, 150 WSS; TTS up to 45; Voice Agent up to 45
Pay As You GoFree
  • All public endpoints
  • Community support
Growth$200 credit then pay-as-you-go ($4K+/year)
  • Higher concurrency limits
  • Standard uptime SLA
EnterpriseContact Sales
  • Custom volumes
  • SLAs
  • Dedicated support

// security & compliance

SOC 2 ISO 27001 GDPR HIPAA

// features

  • Unified Voice Agent API (STT + TTS + LLM orchestration)
  • Flux multilingual conversational STT with turn detection and interruption handling
  • Nova-3 models for high-accuracy transcription in 45+ languages
  • Audio Intelligence add-ons (Redaction, Keyterm Prompting, Smart Formatting)
  • Custom model training
  • Self-hosted and cloud deployment options

sources