Skip to content
API Rates Per-unit API pricing — verified, dated, and logged when it changes
Speech APIs (TTS & STT)
ELEVENLABS, DEEPGRAM

ElevenLabs’ premium TTS beats Deepgram’s modest rates, but Deepgram still wins on raw STT cost

Published Pricing verified

When you need a voice that sounds like a human narrator and a transcription engine that barely nudges the meter, the price tags on ElevenLabs and Deepgram split the difference in a way that matters to product teams. ElevenLabs charges a premium for its high‑fidelity text‑to‑speech models, while Deepgram undercuts it on both pre‑recorded and streaming speech‑to‑text, making the choice hinge on which capability you value most.

### Where the dollars land per operation
| Vendor | Service | Rate |
|--------|---------|------|
| ElevenLabs | Text to Speech (v3) | $0.1 per 1K characters |
| ElevenLabs | Text to Speech (Flash/Turbo) | $0.05 per 1K characters |
| ElevenLabs | Speech to Text (Scribe v2) | $0.22 per hour |
| ElevenLabs | Speech to Text (Scribe v2 Realtime) | $0.39 per hour |
| Deepgram | Speech to Text (Nova‑3, pre‑recorded, monolingual) | $0.0043 per minute |
| Deepgram | Speech to Text (Nova‑3, streaming, monolingual) | $0.0048 per minute |
| Deepgram | Text to Speech (Aura‑2) | $0.03 per 1K characters |

ElevenLabs’ standard TTS model costs $0.1 per thousand characters, roughly double Deepgram’s Aura‑2 rate of $0.03 per thousand characters. The Flash/Turbo variant halves that to $0.05, still above Deepgram’s offering but promising faster turnaround for high‑volume pipelines. On the transcription side, ElevenLabs bills by the hour—$0.22 for batch processing and $0.39 for realtime—whereas Deepgram measures by the minute, with pre‑recorded audio at $0.0043 and streaming at $0.0048. Converting minutes to hours shows Deepgram’s per‑hour cost stays well under $0.30, giving it a clear price advantage for continuous audio streams.

### Who each service is built for
If your product demands studio‑grade voiceovers, the slight premium on ElevenLabs’ TTS is justified by its richer prosody and the option to choose between the full‑featured v3 and the speed‑focused Flash/Turbo. Companies that prioritize brand‑consistent narration—podcast generators, e‑learning platforms, or marketing video tools—will likely lean toward ElevenLabs despite the higher per‑character charge. Conversely, developers whose primary pain point is cheap, reliable transcription will find Deepgram’s minute‑based pricing irresistible. The Nova‑3 engine is engineered for monolingual workloads, making it ideal for call‑center analytics, subtitle generation, or any scenario where volume trumps multilingual flexibility.

In practice, a startup building a voice‑assistant might pair Deepgram’s low‑cost STT for user utterances with ElevenLabs’ Flash/Turbo TTS for responses, balancing cost and quality. An enterprise that needs high‑fidelity narration at scale may accept the $0.1 per 1K characters to avoid the post‑production polishing that cheaper TTS often requires. The decision ultimately rests on whether you are paying for the sound of the voice or the silence of the transcript.

Pricing verified on 2026-09-12.

Common questions

Is there a free tier for either ElevenLabs or Deepgram?

The current pricing tables list only paid rates; no free tier is mentioned.

How is ElevenLabs’ speech‑to‑text billed?

ElevenLabs bills speech‑to‑text by the hour at $0.22 per hour for batch and $0.39 per hour for realtime.

Which vendor offers the cheapest text‑to‑speech rate?

Deepgram’s Aura‑2 costs $0.03 per 1K characters, cheaper than ElevenLabs’ Flash/Turbo at $0.05 and v3 at $0.1.

Sources
  1. ElevenLabs — pricing
  2. Deepgram — pricing