Azure AI Fundamentals (AI-901) · Free practice question 12 of 12
Speaker diarization in transcripts
Hartwell Clinic records two-person telehealth consultations and wants each transcript to show which participant said each sentence. Which Azure Speech capability provides this?
- A.Speaker diarization during speech to text transcription
- B.Pronunciation assessment
- C.Neural text to speech with SSML
- D.Key phrase extraction in Azure Language
Show answer and explanation
Correct answer: A. Speaker diarization during speech to text transcription
Why: Diarization separates the voices in an audio recording and labels each recognized phrase with a speaker, so the transcript shows who said what. Pronunciation assessment scores how accurately a speaker pronounces words, text to speech generates audio rather than transcribing it, and key phrase extraction finds the main topics in text without identifying speakers.
More free Azure AI Fundamentals (AI-901) questions
- Entity linking to disambiguate mentions
- Pronunciation assessment for learners
- Spoken language identification
- Text-to-speech avatar videos
- Custom NER for domain entity types
- Text analytics for health on clinical notes
- Field schema suggestion for new documents
- Base analyzer for custom document analyzers
- Publishing agents to Teams and Copilot
- Generating video from text prompts
- Agent-only guardrail intervention points