Open API PricingOpen API Pricing

Category

Speech-to-Text APIs pricing

Speech-to-Text APIs convert audio into text, spanning batch (async file) and real-time streaming transcription, with add-ons like speaker diarization, translation, and PII redaction. The category splits into focused voice-AI specialists (Deepgram, AssemblyAI, Speechmatics, Gladia, Rev AI) optimized for accuracy, latency, and generous self-serve free tiers, versus hyperscaler platforms (Google, AWS) and the model-API generalist (OpenAI Whisper) that ride massive infrastructure but offer thinner DX and stingier free tiers. Compare on documentation/DX quality, reliability and proven scale, SDK breadth and ecosystem, and how fast a developer or AI agent can self-serve a working key against transparent public pricing.

APIBilled byFree tierVerified
Deepgram
Speech-to-Text APIs
Per minute (usage-based)Free tier or trialJun 27, 2026
AssemblyAI
Speech-to-Text APIs
Per minute (usage-based)Free tier or trialJun 27, 2026
Per minute (per-second billed)No free tierJun 27, 2026
Google Cloud Speech-to-Text
Speech-to-Text APIs
Per 15 secondsFree tier or trialJun 27, 2026
Amazon Transcribe
Speech-to-Text APIs
Per minute (volume-tiered)Free tier or trialJun 27, 2026
Speechmatics
Speech-to-Text APIs
Per minute (usage-based)Free tier or trialJun 27, 2026
Gladia
Speech-to-Text APIs
Per hour (usage-based)Free tier or trialJun 27, 2026
Rev AI
Speech-to-Text APIs
Per minute (usage-based)Free tier or trialJun 27, 2026

Compare pricing side by side