Courseiva

Generative AI Leader Google Cloud's Generative AI Offerings Practice Question

A media company wants to generate short video clips from text prompts for social media ads. The creative team needs a managed Google Cloud service that produces video from descriptive prompts and offers controls for aspect ratio and duration, without managing GPU infrastructure. Which offering should they use?

⚠ Common exam trap

The trap here is assuming any Google generative media model produces video, when Imagen is image-only and Veo is the model purpose-built for text-to-video generation.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

Vertex AI Veo

The team needs prompt-driven video generation delivered as a managed service, which is what Veo on Vertex AI provides, including controls over clip attributes and no infrastructure management. Imagen produces still images, Chirp targets speech, and Cloud Vision API analyzes existing images, so none of the alternatives can generate video content for the social ad campaign.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    Vertex AI Chirp

    Why it's wrong here

    Chirp refers to Google's speech-focused models used for speech-to-text and text-to-speech capabilities. These address audio transcription and synthesis rather than generating visual video content from prompts. A media team needing motion clips for social campaigns would gain nothing from speech models, so Chirp does not fulfill the described generative video requirement even though it is part of the Vertex AI model portfolio.

  • ✓

    Vertex AI Veo

    Why this is correct

    Veo is Google's generative video model available through Vertex AI, capable of producing short video clips from text prompts with controls over characteristics such as aspect ratio and duration. Because it is offered as a managed model on Vertex AI, the creative team avoids provisioning or managing GPUs. This directly matches the requirement to generate video from descriptive prompts in a managed fashion.

  • ✗

    Vertex AI Imagen

    Why it's wrong here

    Imagen is a generative image model that creates and edits still images from text prompts. It does not produce video clips, so it cannot satisfy a requirement for motion content for social ads. While Imagen is also a managed Vertex AI model with strong prompt adherence for images, the deliverable here is video, which places it outside the scenario despite being a legitimate generative media offering.

  • ✗

    Cloud Vision API

    Why it's wrong here

    Cloud Vision API performs analysis tasks such as label detection, OCR, and moderation on existing images. It is a discriminative analysis service, not a generative model, so it cannot create video from text prompts. Deploying it would leave the creative team without any way to synthesize new clips, making it an incorrect choice for this generative video use case.

About these practice questions

This Generative AI Leader question is part of Courseiva's 1,008-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official Google Cloud exam blueprint

This Generative AI Leader practice question is part of Courseiva's free Google Cloud certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the Generative AI Leader exam.