New Universal-3.5 Pro is here. Learn more: Async Realtime
Features

Speaker Identification

Go beyond “Speaker A” and “Speaker B.” Replace generic labels with real names or roles — inferred from the conversation itself, no voice enrollment needed.

Get started with less than 10 lines of code

Turn on speaker labels and Speaker Identification in your transcription request, and get a transcript where each speaker is labeled by name or role.

Identify by name or role

Provide names when you know them, or use roles like Agent, Customer, or Host when you know the function but not the person.

No voice enrollment needed

The model infers identities from conversational context and names mentioned in the audio — you can add metadata to sharpen accuracy.

Pairs with diarization

Diarization separates who spoke when; identification assigns meaningful labels to those speakers — our widest add-on pairing.

Use cases

Know who said what

Attribute every line to a real person or role so transcripts read like conversations.

Attribute meeting notes to people

Label agent and customer

Measure talk time by speaker

Credit podcast hosts and guests

Flag compliance by speaker

Build searchable interviews

Power role-aware analytics

Clean up multi-speaker transcripts

Join 200K+ developers building new experiences with voice data

AssemblyAI's managed API endpoint and diarization won me over — something Whisper couldn't provide.

Josh Mohrer

Josh Mohrer

Founder, Wave.co

If you have an hour of content, the difference between 99% accuracy and 97% accuracy, it's a lot of time for that person to review. So you could cut down their workflow from taking half an hour, to 20 minutes, to 15 minutes — it's huge, right?

Joshua Grossberg

Joshua Grossberg

CTO, Kapwing

Investments in STT improvements always pay for themselves, since it is such a critical building block of the voice pipeline.

Lindsay Liu

Lindsay Liu

Co-Founder & CEO, Super

We needed a provider that could scale with us — offering unlimited concurrent streams, fair pricing, and responsive support.

Mark Barbir

Mark Barbir

CEO, Earmark

The transcription accuracy, reliability, and speed of AssemblyAI's API have greatly enhanced our operations.

Raj Shankar

Raj Shankar

SVP Product, Calabrio

Our free to paid conversion rate doubled after implementing AssemblyAI.

Colin Treseler

Colin Treseler

Founder & CEO, Supernormal

The accuracy was strong, but the great documentation and unique models like Auto Chapters and Sentiment Analysis is what really won us over.

Nathan Webb

Nathan Webb

Product Manager, Aloware

Calls are and will remain a pertinent part of the customer service journey. Customer service isn't moving entirely to chatbots or chat interactions.

Dr. Shane Lynn

Dr. Shane Lynn

CEO, EdgeTier