nerdexam
Microsoft

AI-900 · Question #232

You have 100 instructional videos that do NOT contain any audio. Each instructional video has a script. You need to generate a narration audio file for each video based on the script. Which type of…

The correct answer is C. speech synthesis. Text to speech enables your applications, tools, or devices to convert text into humanlike synthesized speech. The text to speech capability is also known as speech synthesis. Use humanlike prebuilt neural voices out of the box, or create a custom neural voice that's unique to…

Submitted by devops_kid· Mar 30, 2026[ERROR: No official exam domains were provided in the prompt]

Question

You have 100 instructional videos that do NOT contain any audio. Each instructional video has a script. You need to generate a narration audio file for each video based on the script. Which type of workload should you use?

Options

  • Alanguage modeling
  • Bspeech recognition
  • Cspeech synthesis
  • Dtranslation

How the community answered

(55 responses)
  • A
    2% (1)
  • B
    4% (2)
  • C
    84% (46)
  • D
    11% (6)

Explanation

Text to speech enables your applications, tools, or devices to convert text into humanlike synthesized speech. The text to speech capability is also known as speech synthesis. Use humanlike prebuilt neural voices out of the box, or create a custom neural voice that's unique to your product or brand. https://learn.microsoft.com/en-us/azure/cognitive-services/speech-service/text-to-speech

Topics

#Speech synthesis#Text-to-speech#AI workloads

Community Discussion

No community discussion yet for this question.

Full AI-900 Practice