AI-102 Implement computer vision solutions Practice Question
Which TWO Azure AI services can be used to extract text from images?
⚠ Common exam trap
Many exam-takers confuse Azure AI Video Indexer's ability to extract text from video frames as a primary image text extraction service, but it is designed for video analysis and indexing, not standalone image text extraction.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Azure AI Document Intelligence
Azure AI Document Intelligence (formerly Form Recognizer) includes the Read OCR engine that extracts printed and handwritten text from images and documents. Azure AI Computer Vision provides the OCR API (optical character recognition) which can extract text from images, including both printed and handwritten text, and supports multiple languages. Both services are designed specifically for text extraction from visual content.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
Azure AI Face
Why it's wrong here
Azure AI Face detects and identifies human faces, returning bounding boxes and attributes such as age or emotion; it performs no optical character recognition, so it cannot extract text from images. It is tempting because it is an image-processing service, and it would be correct for scenarios such as facial recognition, identity verification or detecting people in photos.
- ✓
Azure AI Document Intelligence
Why this is correct
Azure AI Document Intelligence performs OCR plus layout analysis, extracting text from images embedded in documents and returning structured content. This satisfies the requirement to extract text from images, particularly where layout and form structure matter.
- ✗
Azure AI Video Indexer
Why it's wrong here
Azure AI Video Indexer analyses video and audio streams, producing transcripts, speakers and topics; it does not perform OCR on still images. It is tempting because it does extract text-like content from media, and it would be correct for indexing a video library, generating captions or searching spoken words across recordings.
- ✓
Azure AI Computer Vision
Why this is correct
Azure AI Computer Vision provides the Read and OCR capabilities that extract printed and handwritten text from images, satisfying the requirement to pull text from image files. Its optical character recognition directly returns extracted text without needing document-level structure.
- ✗
Azure AI Custom Vision
Why it's wrong here
Azure AI Custom Vision trains image classification and object detection models on your labelled images; it returns predicted tags or bounding boxes, not recognised characters. It is tempting because it processes images, and it would be correct for scenarios such as detecting product defects, identifying logos or classifying photographs into custom categories.
Go deeper
Related to this question
About these practice questions
One of 761 original AI-102 practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →
JA
Written by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
This AI-102 practice question is part of Courseiva's free Microsoft certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the AI-102 exam.