easyMultiple Choice
Generative AI Leader Practice Question: A developer wants to generate high-quality images…
A developer wants to generate high-quality images from text descriptions using Google Cloud. Which service should they use?
⚠ Common exam trap
Candidates may confuse Gemini's multimodal capabilities with Imagen's specialized text-to-image generation, but Imagen is the correct service for high-quality image generation from text.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Imagen
Imagen is Google Cloud's text-to-image diffusion model, specifically designed to generate high-quality images from natural language descriptions. It leverages deep learning to produce photorealistic outputs, making it the correct choice for this use case.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
Chirp
Why it's wrong here
Chirp is Google Cloud's speech model family, converting audio to text and supporting speech synthesis. It is tempting because it is a generative AI service, but it operates on audio rather than pixels; text-to-image generation requires Imagen, which synthesises images from prompts.
- ✓
Imagen
Why this is correct
Imagen is Google Cloud's dedicated text-to-image generation model, producing high-quality visuals directly from text prompts. It satisfies the stem's requirement for image generation, unlike Gemini's multimodal focus or Vertex AI's broader platform role. The developer should select Imagen as the purpose-built service for this task.
- ✗
Codey
Why it's wrong here
Codey is Google Cloud's code-generation model family, producing or completing source code rather than images. It is tempting because it is a Vertex AI generative model, but text-to-image generation requires Imagen, which is purpose-built for synthesising images from text prompts.
- ✗
Gemini
Why it's wrong here
Gemini is a multimodal model for text, reasoning and chat tasks, not Google Cloud's dedicated text-to-image generator. It is tempting because it accepts image inputs and can describe visuals, but producing high-quality images from text descriptions requires Imagen on Vertex AI.
Go deeper
Related to this question
About these practice questions
One of 1,008 original Generative AI Leader practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →
JA
Written by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
This Generative AI Leader practice question is part of Courseiva's free Google Cloud certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the Generative AI Leader exam.