Microsoft Certified: Azure AI Engineer AssociateImplement natural language processing solutionsEasy

A developer is creating an interactive voice assistant for a smart home system. The assistant needs to respond to user commands like 'Turn on the living room lights' or 'Set the thermostat to 72 degrees'. The key requirement is to generate natural-sounding, human-like speech responses to confirm actions or provide information. Which Azure AI Speech feature should the developer use?

  1. AAzure AI Speech - Neural Text-to-Speech (TTS)
  2. BAzure AI Speech - Speaker Diarization
  3. CAzure AI Speech - Custom Speech
  4. DAzure AI Speech - Speech-to-text
Show answer & explanation

Correct answer: A. Azure AI Speech - Neural Text-to-Speech (TTS)

Neural Text-to-Speech (TTS) is specifically designed to convert text into highly natural-sounding, human-like speech. This is essential for interactive voice assistants to provide engaging and clear verbal responses to users.

Why the other options are wrong

  • B. Speaker Diarization identifies different speakers, not for generating speech.
  • C. Custom Speech is for improving speech *recognition* accuracy, not for generating speech.
  • D. Speech-to-text converts audio to text, which is the opposite of generating speech responses.

Azure AI Speech Neural Text-to-Speech (TTS)

A feature of Azure AI Speech that converts written text into highly natural-sounding, human-like synthetic speech using deep neural networks, offering various voices, styles, and languages.

  • Generates natural and expressive speech from text.
  • Supports many languages and regional accents.
  • Offers a wide selection of pre-built neural voices.
  • Can be customized for specific vocal styles.

Memory trick: Type the words, Hear the voice.

More Implement natural language processing solutions questions