Knowledge · 00:02:38 · 366 words
file
who-said-what.reading
speakers
vocateca (reading)
address
vocateca.com/en/knowledge/who-said-what
duration
00:02:38 · 366 words
Knowledge · Foundations

Who said what

Attribution
00:00:00vocateca

A sentence in a transcript that nobody said in particular is not a fact you can use. Who said it is part of what was said - diarization is what puts a name to a line.

00:00:16

A claim needs an author

00:00:16vocateca

Take the speaker out of a transcript and a claim turns into a rumor - present in the text, attached to no one. It reads the same whether the host said it, a guest said it, or someone read it off a slide.

00:00:34vocateca

That is a real gap, not a cosmetic one. Who said something changes what it is worth: an offhand line from a guest is not the same fact as a considered claim from the host, even if the words on the page are identical.

00:00:53

Diarization labels who spoke when

00:00:53vocateca

Vocateca runs speaker diarization on-device, on by default. It separates the recording into speaker turns and labels them in the output, so the transcript shows who said each line, not just what the line was.

00:01:08vocateca

That labeling happens at the same time as the transcription, not as a separate pass you have to remember to run. A two-person interview or a panel with five voices comes out with turns already split by speaker.

00:01:24

Sourcing a quote to a person

00:01:24vocateca

Once the speaker labels exist, a quote can be attributed to whoever said it, not just to the episode it came from. That distinction matters the moment you are citing a claim in a post or checking who actually made an argument in a recording with more than one voice.

00:01:46vocateca

An unlabeled transcript can tell you what was said in a show. A labeled one can tell you what a specific person said, which is the version you can actually cite.

00:01:59

Labeled, not perfect

00:01:59vocateca

Diarization separates voices and assigns labels; it does not guarantee every label is right on every line, especially with overlapping speech or a voice that only appears for a sentence. Treat labels as a strong default, not a claim of perfect accuracy.

00:02:17vocateca

Even with that caveat, a transcript with speaker turns beats one without them. Attributable and mostly right is more useful than anonymous and technically complete.

00:02:28

Sources

[This article cites no sources. When an article does, they appear here as numbered notes, like on the comparison pages.]

End of reading · 00:02:38

Read on

All articles

Download for macOSFree · macOS 15 or newer, Apple silicon
Version 2.2.5 · 28 August 2026