Skip to main content
mikejaz2
Inspiring
July 26, 2026
Question

Why does transcription confuse voices?

  • July 26, 2026
  • 2 replies
  • 7 views

(Unbelievable...THIRD time I’ve been thru this process of editing my question and the retyping my context. Who designed this dialog, the Department of Redundancy Department?)

 

Anyway, I transcribed a sequence, and Transcription can’t distinguish between a thirty-something, deep voiced person and an older thin voiced male. Trust me, their voices are different.

Is there a way to “tune” the AI so it figures out that there are two speakers here? Right now, the transcription lists any speaker as “Speaker 2”

    2 replies

    mikejaz2
    mikejaz2Author
    Inspiring
    July 27, 2026

    Hey, thanks for the reply...I wasn't even sure if my query was posted. Kinda glitchy there.

    Anyway, the audio ain't great, but it certainly is legible. There's two camera mic tracks, and a Zoom that was somewhere near the subject. I switched from "mixed" audio and tried each of the tracks individually,  and finally got a transcript that distinguished between the two speakers properly.

    Again, thanks.

    Community Expert
    July 27, 2026

    Does the audio need to be cleaned up or enhanced in any way?