← Back to blog

Automatic diarization: separating voices in audio

Diarization (or speaker separation) is a technology that automatically identifies the different speakers in an audio recording. It's an essential feature for meeting minutes, interviews and conferences.

What is diarization?

Diarization is the process that answers the question 'who spoke when?'. The algorithm analyzes each speaker's vocal characteristics (timbre, frequency, rhythm) to tell them apart and assign each speech segment to the right person.

How to use diarization in Voqia

In Voqia, enable the 'Separate speakers' option before importing your audio file. Our AI analyzes the recording and identifies each speaker. You can then rename speakers (e.g. 'Speaker 1' -> 'Marie') for clear minutes.

Diarization use cases

Diarization is especially useful for team meetings (identifying who made which decision), journalistic interviews (telling the interviewer from the interviewee), medical consultations (separating doctor from patient), and conferences with multiple speakers.

Diarization software: the criteria that matter

Before choosing diarization software, check three things: how many speakers it handles (Voqia tells voices apart automatically without declaring them upfront), whether you can rename speakers afterwards, and the price — many tools lock diarization behind paid plans, while it's included for free in Voqia, both live and on files.

Diarization turns a raw transcript into a structured, usable document. With Voqia, this advanced feature is available for free.

Try Voqia for free

10 free transcription minutes, voice cloning and text-to-speech included. No credit card.

Get started free
← Back to blog