Microsoft Certified: Azure AI Engineer AssociateImplement natural language processing solutionsEasy

A developer wants to integrate a feature into their application that converts text into natural-sounding speech. The primary goal is to provide a high-quality, human-like voice experience for users, without needing a custom-branded voice. The solution must support multiple languages. Which Azure AI Speech feature should the developer use?

  1. ACustom Neural Voice
  2. BSpeech Translation
  3. CNeural Text-to-Speech (TTS)
  4. DBatch Speech-to-Text
Show answer & explanation

Correct answer: C. Neural Text-to-Speech (TTS)

Neural Text-to-Speech (TTS) is the Azure AI Speech feature designed to convert text into highly natural-sounding, human-like speech using pre-built neural voices. It supports multiple languages and is ideal when a custom-branded voice is not required. Batch Speech-to-Text converts audio to text, Custom Neural Voice creates unique voices, and Speech Translation translates spoken language.

Why the other options are wrong

  • A. Custom Neural Voice creates a unique, branded voice, but the requirement specifically states 'without needing a custom-branded voice'.
  • B. Speech Translation translates spoken language, not converts text to speech directly.
  • D. Batch Speech-to-Text converts audio to text, which is the opposite of the requirement.

Neural Text-to-Speech (TTS)

An Azure AI Speech service that synthesizes highly natural-sounding speech from text using deep neural networks and pre-built voices.

  • Generates human-like and expressive speech.
  • Supports numerous languages and voice styles.
  • Does not require training with custom voice data.

Memory trick: Neural TTS makes text talk naturally, like a digital narrator.

More Implement natural language processing solutions questions