Skip to main content
mikejaz2
Inspiring
July 26, 2026
Answered

Why does transcription confuse voices?

  • July 26, 2026
  • 3 replies
  • 18 views

(Unbelievable...THIRD time I’ve been thru this process of editing my question and the retyping my context. Who designed this dialog, the Department of Redundancy Department?)

 

Anyway, I transcribed a sequence, and Transcription can’t distinguish between a thirty-something, deep voiced person and an older thin voiced male. Trust me, their voices are different.

Is there a way to “tune” the AI so it figures out that there are two speakers here? Right now, the transcription lists any speaker as “Speaker 2”

    Correct answer mikejaz2

    Hey, thanks for the reply...I wasn't even sure if my query was posted. Kinda glitchy there.

    Anyway, the audio ain't great, but it certainly is legible. There's two camera mic tracks, and a Zoom that was somewhere near the subject. I switched from "mixed" audio and tried each of the tracks individually,  and finally got a transcript that distinguished between the two speakers properly.

    Again, thanks.

    3 replies

    mikejaz2
    mikejaz2AuthorCorrect answer
    Inspiring
    July 27, 2026

    Hey, thanks for the reply...I wasn't even sure if my query was posted. Kinda glitchy there.

    Anyway, the audio ain't great, but it certainly is legible. There's two camera mic tracks, and a Zoom that was somewhere near the subject. I switched from "mixed" audio and tried each of the tracks individually,  and finally got a transcript that distinguished between the two speakers properly.

    Again, thanks.

    caroline_edits
    Community Manager
    Community Manager
    July 27, 2026

    Glad to hear you found a way to get the different voices identified. Thanks for taking the time to share what worked for you!

    Caroline

    Community Expert
    July 27, 2026

    Does the audio need to be cleaned up or enhanced in any way?