Microsoft Certified: Azure AI Engineer AssociateImplement natural language processing solutionsMedium
A contact center wants to analyze recorded customer service calls to identify recurring themes, common customer issues, and agent performance. The calls are long, and manually listening to all of them is impractical. The solution needs to process these recorded audio files in bulk to extract insights. Which Azure AI service is the primary component for converting these audio recordings into text for further analysis?
- AAzure AI Vision for object detection
- BAzure AI Language for PII detection
- CAzure AI Speech for real-time speech-to-text
- DAzure AI Speech for batch speech-to-text
Show answer & explanationAnswer & explanation
Correct answer: D. Azure AI Speech for batch speech-to-text
To process long, recorded audio files in bulk, Azure AI Speech's batch speech-to-text service is the most appropriate choice. Real-time transcription is for live audio streams, while batch is for pre-recorded files, which fits the scenario of analyzing recorded customer calls.
Why the other options are wrong
- A. Object detection is for images/video, not audio transcription.
- B. PII detection works on text, not directly on audio, and is a subsequent step after transcription.
- C. Real-time speech-to-text is for live audio, not pre-recorded files in bulk.
Batch Speech-to-Text
A feature of Azure AI Speech that allows for asynchronous transcription of large volumes of pre-recorded audio files into text.
- Designed for processing recorded audio.
- Supports various audio formats and languages.
- Ideal for transcribing archives, call recordings, and media content.
Memory trick: Batch processing is like a tireless scribe for all your recordings.