CertKeen

Azure AI Fundamentals (AI-901) · Free practice question 12 of 12

Speaker diarization in transcripts

Hartwell Clinic records two-person telehealth consultations and wants each transcript to show which participant said each sentence. Which Azure Speech capability provides this?

  1. A.Speaker diarization during speech to text transcription
  2. B.Pronunciation assessment
  3. C.Neural text to speech with SSML
  4. D.Key phrase extraction in Azure Language
Show answer and explanation

Correct answer: A. Speaker diarization during speech to text transcription

Why: Diarization separates the voices in an audio recording and labels each recognized phrase with a speaker, so the transcript shows who said what. Pronunciation assessment scores how accurately a speaker pronounces words, text to speech generates audio rather than transcribing it, and key phrase extraction finds the main topics in text without identifying speakers.

More free Azure AI Fundamentals (AI-901) questions