Courseiva

AI-900 Practice Question: Describe features of generative AI workloads on Azure

What is 'Whisper' in Azure OpenAI and what can it do?

⚠ Common exam trap

The trap here is that the name 'Whisper' might mislead candidates into thinking it relates to quiet audio output (text-to-speech) or a low-power mode, when in fact it is a speech recognition model for transcribing audio to text.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

A speech recognition model that transcribes audio files to text across 100+ languages

Whisper is a speech recognition model available in Azure OpenAI that transcribes audio files into text. It supports over 100 languages and is designed for high accuracy in diverse acoustic environments, making it ideal for tasks like meeting transcription, voice note conversion, and multilingual audio processing.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • A low-power mode for running Azure OpenAI at reduced compute cost

    Why it's wrong here

    This would refer to Azure cost optimization features like reserved capacity, serverless deployments, or dev/test pricing, which are operational settings for compute infrastructure. Whisper is not an infrastructure mode; it is a specific pre-trained model designed for speech-to-text. Reducing compute cost would not change the model's core function of transcribing audio.

  • A speech recognition model that transcribes audio files to text across 100+ languages

    Why this is correct

    This correctly identifies Whisper as an automatic speech recognition model that converts spoken language in audio files into written text. It supports over 100 languages and can also translate non-English speech into English text. Whisper is available through Azure OpenAI for batch transcription tasks on pre-recorded content, not for real-time conversational streaming.

  • A secure communication channel for transmitting sensitive data to Azure OpenAI

    Why it's wrong here

    This answer describes Azure networking capabilities such as Azure Private Link or service endpoints that encrypt and isolate data in transit, but these are infrastructure security controls, not an AI model. Whisper is a speech recognition model offered through Azure OpenAI, not a communication channel. The error is confusing the underlying data-plane security with the actual model used for audio transcription.

  • A text-to-speech model that generates very quiet, whispered audio output

    Why it's wrong here

    This is a pun on the model's name: Whisper has nothing to do with low volume or whispered output. In fact, Whisper is an automatic speech recognition (ASR) model that takes audio input and produces text transcriptions, the exact opposite of a text-to-speech (TTS) synthesizer. It cannot generate audio, quiet or otherwise; it consumes audio and outputs text.

About these practice questions

This AI-900 question is part of Courseiva's 985-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

This AI-900 practice question is part of Courseiva's free Microsoft certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the AI-900 exam.