New Universal-3.5 Pro is here. Learn more: Async Realtime
Biography Podcast

Build Voice AI apps on industry-leading speech models

We're giving Biography podcast listeners $100 in free credits to try AssemblyAI's speech-to-text and Speech Understanding models. We can't wait to see what you build.

Speech-to-text

Unlock the value of pre-recorded voice data and power workflows with unmatched accuracy.

Streaming speech-to-text

Build voice agent workflows with ultra-low latency, high accuracy, and precise end-of-turn control.

Speech Understanding

Enable deep analysis and high-value insights with sophisticated Speech Understanding models.

Metaview
Ashby
Cluely
Genio
Siro

36%

improvement in close rate

See case study
LiveKit
Earmark
Commure
Dovetail
Fireflies

“The new Universal-3.5 Pro speech model from AssemblyAI is best so far in terms of accuracy, latency, and language switching.”

Retell
CallRail
Apollo.io
ClickUp
Calabrio

80%

increase in customer satisfaction

HeyGen
Granola
Siro
JotPsych
Granola

“Assembly has saved us countless hours managing models, and provided exceptional accuracy.”

Metaview
Ashby
Cluely
Genio
Siro

36%

improvement in close rate

See case study
LiveKit
Earmark
Commure
Dovetail
Fireflies

“The new Universal-3.5 Pro speech model from AssemblyAI is best so far in terms of accuracy, latency, and language switching.”

Retell
CallRail
Apollo.io
ClickUp
Calabrio

80%

increase in customer satisfaction

HeyGen
Granola
Siro
JotPsych
Granola

“Assembly has saved us countless hours managing models, and provided exceptional accuracy.”

Quickstart

Start building with credits on us

Create a free account, drop in your API key, and transcribe your first file in minutes. Test in the no-code playground, then copy a ready-made request into your app—the credits are already waiting.

Claim $100 in free credits

No credit card required

Word error rate

Your product is only as good as the inputs it's built on. Word error rate is the share of words the model gets wrong against a human reference—the standard measure of transcription accuracy.

Pre-recorded word error rate on English audio.

*Lower is better*
AssemblyAI Universal-3 Pro
4.50%
Mistral Voxtral Mini
5.24%
OpenAI GPT-4o Transcribe
5.34%
Deepgram Nova-3
6.66%
Azure Batch
7.02%

Source: AssemblyAI published benchmarks — assemblyai.com/benchmarks.