mediumMultiple Choice
Generative AI Leader Practice Question: A research team wants to use Google's AI to…
A research team wants to use Google's AI to generate video content from text prompts for a creative project. Which Google Cloud generative AI model should they use?
⚠ Common exam trap
Generative AI Leader often tests model-to-modality mapping — candidates confuse Imagen (image) with Veo (video) because both are generative media models with similar-sounding names.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Veo
Veo is Google's generative AI model specifically designed for text-to-video generation, making it the correct choice for creating video content from text prompts. It is part of Google's Vertex AI model portfolio and supports high-definition video generation with cinematic controls. Imagen, Codey, and Gemini serve different modalities.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
Imagen
Why it's wrong here
Imagen generates and edits images from text prompts, not video. It is tempting because it is Google Cloud's flagship text-to-media model, and would be correct for producing still imagery, but the scenario requires video output, which Imagen cannot produce.
- ✗
Codey
Why it's wrong here
Codey specialises in code completion and generation from natural-language prompts, producing source code rather than video. It is tempting because it is a Google Cloud generative model, and would be correct for software development tasks, but it cannot synthesise video content.
- ✓
Veo
Why this is correct
Veo is Google Cloud's generative video model, accepting text prompts and producing video clips, which directly satisfies the stem's requirement to generate video content from text. Other Gemini and Imagen models output text or images respectively, so they cannot fulfil the video-generation constraint.
- ✗
Gemini
Why it's wrong here
Gemini is a multimodal model handling text, images, audio and code, but it does not generate video files. It is tempting as Google's most capable general model, and suits reasoning or content drafting, yet text-to-video generation requires Veo.
Go deeper
Related to this question
About these practice questions
One of 1,008 original Generative AI Leader practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →
JA
Written and reviewed by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
Last reviewed September 2026 · checked against the official Google Cloud exam blueprint
This Generative AI Leader practice question is part of Courseiva's free Google Cloud certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the Generative AI Leader exam.