Generative AI Leader Google Cloud's Generative AI Offerings Practice Question
A media company wants to generate short video clips from text prompts for social media ads. The creative team needs a managed Google Cloud service that produces video from descriptive prompts and offers controls for aspect ratio and duration, without managing GPU infrastructure. Which offering should they use?
⚠ Common exam trap
The trap here is assuming any Google generative media model produces video, when Imagen is image-only and Veo is the model purpose-built for text-to-video generation.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Vertex AI Veo
The team needs prompt-driven video generation delivered as a managed service, which is what Veo on Vertex AI provides, including controls over clip attributes and no infrastructure management. Imagen produces still images, Chirp targets speech, and Cloud Vision API analyzes existing images, so none of the alternatives can generate video content for the social ad campaign.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
Vertex AI Chirp
Why it's wrong here
Chirp refers to Google's speech-focused models used for speech-to-text and text-to-speech capabilities. These address audio transcription and synthesis rather than generating visual video content from prompts. A media team needing motion clips for social campaigns would gain nothing from speech models, so Chirp does not fulfill the described generative video requirement even though it is part of the Vertex AI model portfolio.
- ✓
Vertex AI Veo
Why this is correct
Veo is Google's generative video model available through Vertex AI, capable of producing short video clips from text prompts with controls over characteristics such as aspect ratio and duration. Because it is offered as a managed model on Vertex AI, the creative team avoids provisioning or managing GPUs. This directly matches the requirement to generate video from descriptive prompts in a managed fashion.
- ✗
Vertex AI Imagen
Why it's wrong here
Imagen is a generative image model that creates and edits still images from text prompts. It does not produce video clips, so it cannot satisfy a requirement for motion content for social ads. While Imagen is also a managed Vertex AI model with strong prompt adherence for images, the deliverable here is video, which places it outside the scenario despite being a legitimate generative media offering.
- ✗
Cloud Vision API
Why it's wrong here
Cloud Vision API performs analysis tasks such as label detection, OCR, and moderation on existing images. It is a discriminative analysis service, not a generative model, so it cannot create video from text prompts. Deploying it would leave the creative team without any way to synthesize new clips, making it an incorrect choice for this generative video use case.
Go deeper
Related to this question
About these practice questions
This Generative AI Leader question is part of Courseiva's 1,008-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →
JA
Written and reviewed by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
Last reviewed September 2026 · checked against the official Google Cloud exam blueprint
This Generative AI Leader practice question is part of Courseiva's free Google Cloud certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the Generative AI Leader exam.