Courseiva
mediumMultiple Choice

Generative AI Leader Practice Question: A research team wants to use Google's AI to…

A research team wants to use Google's AI to generate video content from text prompts for a creative project. Which Google Cloud generative AI model should they use?

⚠ Common exam trap

Generative AI Leader often tests model-to-modality mapping — candidates confuse Imagen (image) with Veo (video) because both are generative media models with similar-sounding names.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

Veo

Veo is Google's generative AI model specifically designed for text-to-video generation, making it the correct choice for creating video content from text prompts. It is part of Google's Vertex AI model portfolio and supports high-definition video generation with cinematic controls. Imagen, Codey, and Gemini serve different modalities.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    Imagen

    Why it's wrong here

    Imagen generates and edits images from text prompts, not video. It is tempting because it is Google Cloud's flagship text-to-media model, and would be correct for producing still imagery, but the scenario requires video output, which Imagen cannot produce.

  • ✗

    Codey

    Why it's wrong here

    Codey specialises in code completion and generation from natural-language prompts, producing source code rather than video. It is tempting because it is a Google Cloud generative model, and would be correct for software development tasks, but it cannot synthesise video content.

  • ✓

    Veo

    Why this is correct

    Veo is Google Cloud's generative video model, accepting text prompts and producing video clips, which directly satisfies the stem's requirement to generate video content from text. Other Gemini and Imagen models output text or images respectively, so they cannot fulfil the video-generation constraint.

  • ✗

    Gemini

    Why it's wrong here

    Gemini is a multimodal model handling text, images, audio and code, but it does not generate video files. It is tempting as Google's most capable general model, and suits reasoning or content drafting, yet text-to-video generation requires Veo.

About these practice questions

One of 1,008 original Generative AI Leader practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official Google Cloud exam blueprint

This Generative AI Leader practice question is part of Courseiva's free Google Cloud certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the Generative AI Leader exam.