Microsoft Certified: Azure AI Engineer AssociateImplement natural language processing solutionsMedium
A developer wants to integrate a feature into their application that converts text into natural-sounding speech. The goal is to provide a variety of realistic voices, including different genders and accents, and to allow for fine-grained control over speech characteristics like pitch and speaking rate. Which Azure AI service feature should be used?
- AAzure AI Speech for Neural Text-to-Speech (TTS)
- BAzure AI Speech for Custom Speech
- CAzure AI Language for Text Summarization
- DAzure OpenAI Service for Generative AI
Show answer & explanationAnswer & explanation
Correct answer: A. Azure AI Speech for Neural Text-to-Speech (TTS)
Neural Text-to-Speech (TTS) in Azure AI Speech is specifically designed to convert text into highly natural-sounding speech with a wide range of pre-built neural voices, supporting different languages, genders, and accents. It also offers SSML (Speech Synthesis Markup Language) for fine-grained control over speech characteristics.
Why the other options are wrong
- B. Custom Speech improves speech *recognition* accuracy, not text-to-speech generation.
- C. Text Summarization condenses text, it does not convert text to speech.
- D. Generative AI creates text content, it does not convert text to speech.
Neural Text-to-Speech (TTS)
A feature of Azure AI Speech that uses deep neural networks to synthesize highly natural and human-like speech from text, offering diverse voices and granular control over output.
- Generates natural, expressive speech.
- Offers a wide variety of pre-built neural voices.
- Supports SSML for pitch, rate, pronunciation control.
- Used for audiobooks, voice assistants, accessibility features.
Memory trick: Neural TTS: Your text finds its perfect voice.