How speaker diarization works — and where it slips
Splitting a recording by who is talking is a different problem from writing down what they said. Here is what the model actually listens for.
2 min read
Notes on transcription, speech AI and getting usable text out of audio.
Splitting a recording by who is talking is a different problem from writing down what they said. Here is what the model actually listens for.
2 min read
Three levels, three different models. What actually changes between them, and when the slowest one earns its wait.
2 min read
Try it
Paste a link or upload a file and read the transcript in minutes.
You have used today's free guest transcriptions. Create an account or try again tomorrow.
Free $0
Already have an account? Log in
Your request is open
We sent a confirmation to your email. You will get our reply by email and can read it on the request page.
Open the request