Courseiva
easyMultiple Choice

Generative AI Leader Practice Question: A developer wants to use a pre-trained model to…

A developer wants to use a pre-trained model to identify objects in images. Which Google Cloud AI API should they use?

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

Vision AI

Vision AI provides pre-trained models for object detection, image classification, etc. Natural Language AI is for text, Speech-to-Text for audio, and Translation for text translation.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    Speech-to-Text

    Why it's wrong here

    Speech-to-Text transcribes audio into written text; it processes acoustic signals, not pixels, so it cannot identify objects in images. It is tempting as a pre-trained Google Cloud AI API, and it would be correct for converting recorded speech into transcripts, not for image object detection.

  • ✗

    Translation API

    Why it's wrong here

    Translation API converts text between languages; its input is a string of characters, not image data, so object detection cannot be performed. It is tempting as a pre-trained Google Cloud AI API, and it would be correct for localising text, not for identifying objects within images.

  • ✗

    Natural Language AI

    Why it's wrong here

    Natural Language AI performs text classification, entity extraction and sentiment analysis on documents; it accepts no image input, so object detection is impossible. It is tempting because it is a pre-trained Google Cloud AI API, and it would be correct for analysing text, not identifying objects in images.

  • ✓

    Vision AI

    Why this is correct

    Vision AI provides pre-trained models purpose-built for image analysis, including object detection and label recognition, so no custom training is needed. It directly satisfies the developer's requirement to identify objects in images using an existing model, unlike APIs scoped to text, speech or video.

About these practice questions

This Generative AI Leader question is part of Courseiva's 1,008-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

This Generative AI Leader practice question is part of Courseiva's free Google Cloud certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the Generative AI Leader exam.