Be able to map a scenario to the right Azure Vision capability: Read for OCR, liveness for anti-spoofing, Custom Vision for classification/detection, and Docker export for edge deployment. Most important: match the feature to the exact requirement, not the most familiar service.
Start practicing
Implement computer vision solutions — choose a session length
Free · No account required
Domain overview
This domain covers Azure AI Vision, Face, and Custom Vision: image classification, object detection, OCR, spatial analysis, and face detection/verification. Questions are scenario-based, asking you to pick the correct feature, SDK call, or deployment option for a stated requirement such as liveness, handwriting, or on-premises inference.
Exam objectives
Selecting Azure AI Vision Read OCR for printed and handwritten text extraction from images
Enabling Face API liveness detection to reject spoofed photos and videos
Deploying a Custom Vision model as a Docker container for on-premises inference
Recognizing Custom Vision actions: upload/tag images, train, evaluate, and publish/export models
Confusing Face detection (bounding boxes, attributes) with Face identification or verification, which require PersonGroup or LargePersonGroup enrollment
Assuming Read OCR handles only printed text, then choosing a weaker feature for mixed handwritten and printed content
Believing Custom Vision can train a model without labeled images or that it performs speech or text analytics tasks
Click any question to see the full explanation and answer options, or start a focused practice session above.
A retail company uses Azure Computer Vision to analyze customer traffic in stores. They deploy a custom object detection model to count customers and detect occupancy. After deployment, the model consistently underestimates the number of customers during peak hours. The company has retrained the model with more data but the issue persists. What is the most likely cause?
2A company uses Azure Face API to verify employee identities for building access. They need to ensure that only live faces are used, not photos or videos. Which feature should they enable?
3A company uses Azure Custom Vision to build a classifier for defect detection on a manufacturing line. They have labeled images of products with and without defects. Which TWO actions should they take to improve model performance?
4You are a data scientist at a healthcare startup. You have deployed a custom object detection model using Azure Custom Vision to detect tumors in MRI scans. The model was trained on 10,000 labeled scans from a single hospital. After deployment, the model performs well on scans from that hospital but poorly on scans from a different hospital with a different MRI machine. The new hospital's scans have slightly different contrast and resolution. The model's precision drops from 0.92 to 0.65, and recall drops from 0.88 to 0.50. You have access to 500 labeled scans from the new hospital. You need to improve the model's performance on the new hospital's data as quickly as possible with minimal effort. What should you do?
5A retail company uses Azure Computer Vision to analyze customer traffic in stores. They process images from security cameras using the OCR API to detect product labels. Recently, the OCR accuracy has decreased for images with poor lighting. Which pre-processing step should the company implement to improve OCR accuracy?
6Which THREE actions can be performed using the Azure Custom Vision service?
7Refer to the exhibit. An Azure Cognitive Services Computer Vision API call for image captioning is returning only one caption. The developer wants to get three possible captions ranked by confidence. Which parameter should be modified in the request?
8Which TWO Azure AI services can be used to extract text from images and PDFs? (Select two.)
9Which TWO Azure services can be used to perform optical character recognition (OCR) on documents? (Select two.)
10You are developing a solution that uses Azure AI Video Indexer to analyze surveillance videos for suspicious activity. The solution must generate alerts when a person is detected in a restricted area. Which feature should you use?
11Which THREE are valid uses of Azure AI Vision Image Analysis 4.0? (Select three.)
12You need to build a solution that detects whether a person is wearing a hard hat in a construction site image. Which Azure AI service should you use?
13You are developing a solution to detect defects on a manufacturing assembly line using computer vision. The solution must classify images as 'defective' or 'non-defective'. You have a limited set of labeled images (500 per class). Which approach should you recommend?
14You need to build a solution that reads text from images in multiple languages, including Arabic and English, and translates the text into English. The solution must preserve the original layout as much as possible. Which combination of Azure AI services should you use?
15You are developing a mobile app that allows users to take a photo of a product and get information about it. The app must identify the product from the image. Which Azure AI service should you use?
16You are deploying a computer vision model using Azure AI Custom Vision with a small dataset of 200 images per class. The model shows high accuracy on training data but low accuracy on test data. Which action should you take to reduce overfitting?
17Which TWO Azure AI services can be used to perform optical character recognition (OCR) on images? (Choose two.)
18Which THREE factors should you consider when choosing between Azure AI Custom Vision and Azure AI Vision pre-built models for an image classification task? (Choose three.)
19You have an Azure AI Vision resource named MyVisionService. You run the above Azure CLI command and get the keys. Your application uses key1 for authentication. You need to rotate the keys without downtime. What should you do?
20You are developing an application that processes images of handwritten forms. The forms contain checkboxes that may be checked or unchecked. Which Azure AI service should you use to detect the state of the checkboxes?
21A manufacturing company uses Azure AI Custom Vision to detect defects on a production line. The model was trained with 500 images per class and achieves 95% accuracy. After deployment, the model's accuracy drops to 80% due to changes in lighting conditions. What is the most effective first step to improve the model's robustness?
22You need to analyze a video stream from a security camera to count the number of people entering a building. Which Azure AI service is most suitable?
23You are deploying a Custom Vision object detection model to an Azure Container Instance for real-time inference. The model must respond within 500 ms. The default container runs on CPU. What should you do to meet the latency requirement?
24You need to detect if a photo contains adult or racy content. Which Azure AI Computer Vision feature should you use?
25You are using Azure AI Custom Vision to classify images of animals. The training set has 1000 images of cats and 1000 images of dogs. After training, the model performs well on the test set. However, when deployed, it misclassifies images of wolves as dogs. What is the most likely cause?
26Which TWO actions should you take to reduce the latency of an Azure AI Computer Vision OCR call on a large image?
27Which THREE factors are critical to consider when designing a custom vision solution for a manufacturing quality inspection system?
28Which TWO Azure AI services can be used to extract text from images?
29Refer to the exhibit. You are creating an Azure Cognitive Services account using an ARM template snippet. What type of account is being created?
30You are building a computer vision solution to detect defects on a manufacturing assembly line. The solution must process images in real-time with low latency, and you need to choose an Azure service. Which service should you use?
31You are designing a solution to detect brand logos in social media images. The logos vary in size and orientation. You need to achieve high accuracy with minimal false positives. Which approach should you recommend?
32You need to analyze videos stored in Azure Blob Storage to detect objects and generate timestamps. Which Azure service should you use?
33A company uses Azure Computer Vision to moderate user-generated content. The solution must detect adult content and flag it. Which API should you call?
34You are deploying a Custom Vision model to a production environment. The model must handle 100 predictions per second with low latency. Which deployment option should you choose?
35You need to extract handwritten text from scanned forms. Which Azure Computer Vision feature should you use?
36Which TWO Azure services can be used to perform optical character recognition (OCR) on images?
37Which THREE factors should you consider when selecting a pricing tier for Azure Computer Vision in a production environment?
38You are creating a new Custom Vision project with the above JSON. The domainId corresponds to the 'Logo' domain. Which type of model will this project train?
39You are building a solution to detect if a person is wearing a hard hat in construction site images. You have a small dataset of labeled images. Which Azure service should you use?
40You are designing a solution that uses Azure AI Vision to extract text from scanned invoices. The invoices vary in layout and include both printed and handwritten fields. The solution must achieve high accuracy with minimal manual labeling. Which approach should you recommend?
41Your team is building a mobile app that uses Azure Custom Vision to classify plant species. The app must work offline and sync labeled images when connectivity is restored. Which SDK feature should you use?
42A manufacturing company uses Azure Custom Vision to detect defects on an assembly line. The model is deployed to a container on a local edge server. Recently, the model's accuracy dropped. You suspect data drift. What should you do to monitor and retrain the model?
43You are building a solution to automatically tag images uploaded to an Azure Storage blob container using Azure AI Vision. The solution must process images as soon as they are uploaded. Which service should you use to trigger the image analysis?
44Refer to the exhibit. You have trained an object detection model in Azure Custom Vision. The model is published as 'defect-model'. You need to deploy this model to a Docker container for on-premises inference using the Azure IoT Edge runtime. What should you do first?
45You have a computer vision solution that analyzes security camera feeds to detect people and vehicles. The solution uses Azure AI Vision Spatial Analysis. You need to ensure compliance with privacy regulations by blurring detected faces. Which feature should you enable?
46Your application needs to determine whether two photos of the same person are of the same individual, even if they are from different angles. Which Azure AI service should you use?
47You are designing a solution that reads handwritten notes from patient intake forms. The solution must handle various handwriting styles. Which Azure AI capability should you use?
48You have a real-time video processing pipeline using Azure AI Video Indexer. You need to detect when a specific person appears in archived video footage. Which approach minimizes latency and cost?
49Your company wants to moderate user-uploaded images for adult content. Which Azure AI service should you use?
50You are building a mobile app that allows users to take a photo of a product and get detailed information. The app uses Azure AI Custom Vision to classify products. You need to ensure low latency for inference. What should you do?
51You are building a document processing solution that extracts information from invoices. The invoices come in various formats and languages. You need to extract line items, totals, and supplier names. Which THREE services should you combine?
52A company is building a solution to analyze customer reviews images using Azure AI Vision. They need to extract text from images that may contain both printed and handwritten text. Which feature should they use?
53A retail company uses Azure AI Vision to analyze shelf images for inventory management. They notice that the Object Detection model sometimes misses small items. What is the most effective way to improve detection of small objects?
54A company wants to moderate user-generated images for adult content. Which Azure AI Vision feature should they use?
55A company uses the Face API to detect and identify employees for building access. They need to ensure that the system complies with GDPR requirements for biometric data. Which action should they take?
56A company uses the Computer Vision Image Analysis API to generate captions for images. The captions are often too generic. How can they improve the descriptiveness of captions?
57A company wants to extract key-value pairs from scanned invoices using Azure AI. Which service should they use?
58Which TWO Azure AI services can be used to detect objects in images?
59You work for a manufacturing company that uses Azure AI services to automate quality inspection on a production line. You have a Custom Vision object detection model that identifies defects on metal parts. The model was trained on images captured under ideal lighting conditions. However, when deployed in the factory, the model's accuracy drops significantly due to inconsistent lighting and glare. You need to improve the model's robustness without collecting new images from the factory floor. What should you do?
60A company is building a computer vision solution using Azure AI Vision to analyze images of retail shelves. The solution must detect product presence and read expiration dates. Which TWO Azure AI Vision features should be used?
61A financial services company is building a computer vision solution to automatically extract data from scanned checks. The solution must recognize handwritten amounts, printed account numbers, and signature presence. The company has a large dataset of labeled check images. They need high accuracy and the ability to retrain with new data. Which Azure service should they use?
62A logistics company uses Azure AI Vision to analyze images of packages on conveyor belts. They need to detect damaged packages and read tracking numbers. The solution must process high throughput (1000 images per minute) with low latency (<500ms per image). The images are captured by fixed cameras. Which approach should you recommend?
63A university is developing an app for students to take photos of handwritten notes and convert them to digital text. The app must support multiple languages including English and Spanish. The solution should use a pre-built AI service. Which Azure service should you use?
64You are developing an app that analyzes images of restaurant receipts. The app must extract the merchant name, transaction date, and total amount from each receipt. You want to minimize development effort and use a prebuilt Azure AI service. Which service should you use?
65You are building a web application that lets users upload photos of restaurant menus and receive the extracted text. The menus are often photographed at an angle, with uneven lighting and background clutter. You need an Azure AI Vision capability that returns text lines and words with bounding boxes, and you want to minimize development effort. Which Azure AI Vision feature should you use?
66You are developing an application that uses Azure AI Vision to analyze images of products on an assembly line. The application must identify the presence of specific objects, such as screws, bolts, and washers, and return their bounding boxes. You have a limited set of labeled images for each object type. Which Azure service should you use to train a model that meets these requirements?
67You are building an Azure AI Vision solution that analyzes live video from a camera mounted on a delivery truck. The solution must read street signs in real time and return the recognized text with bounding box coordinates. You need to minimize latency and cost. Which Azure AI Vision feature should you use?
68A security company uses Azure AI Face API to analyze surveillance footage. They need to detect faces in low-light images and obtain face bounding boxes, but they do not need to identify individuals. They also want to minimize cost and avoid unnecessary features. Which Face API operation should they call?
69A company uses Azure AI Vision Image Analysis to generate captions for product photos. The solution must return a caption in English and a confidence score for each image. You call the Image Analysis API with the caption feature. The response does not include a confidence score. What should you do to obtain confidence scores for the captions?
70You are designing a solution that uses Azure AI Vision to analyze images uploaded by users. You need to extract text and also generate a descriptive caption for each image. You want to use the Image Analysis API. Which two capabilities should you enable? (Choose two.)
71A company needs to identify and tag products on store shelves using a custom model. They have a large dataset of labeled images with bounding boxes around each product. They want to train a model that can detect multiple products in new images and return their locations. Which Azure service should they use?
72You are designing an Azure AI Vision solution that must detect and extract text from identity documents such as passports and driver's licenses. The solution must also identify the document type and extract key fields like name and date of birth. You need to choose the appropriate Azure AI service and features. (Choose two.)
73A developer needs to build a mobile app that identifies dog breeds from photos. They have a small dataset of labeled images and want to train a custom model with minimal machine learning expertise. Which Azure service should they use?
74You are building a web app that allows users to upload images and receive a list of content tags, such as 'outdoor', 'tree', and 'person'. You want to use a prebuilt Azure AI service and minimize custom development. Which service should you use?
75A media company needs to automatically generate descriptive captions for thousands of archived photographs stored in Azure Blob Storage. The solution must be fully managed and require no model training. Which Azure AI Vision capability should they use?
76A company uses Azure AI Vision to extract text from scanned invoices. They need to preserve the layout information, such as tables and key-value pairs, to automate data entry. Which Azure service should they use?
77A developer is building an Azure AI Vision solution that must detect and locate multiple objects in an image, such as cars, people, and traffic lights. The solution must return bounding boxes for each detected object. Which Azure AI Vision capability should the developer use?
78A security firm wants to analyze live video from cameras at a warehouse gate to count the number of people entering and leaving. They require a ready-to-use service that provides a real-time count and does not require training a custom model. Which Azure AI service should they use?
79You are building a web application that allows users to upload images of restaurant receipts and extract the total amount and merchant name. The receipts may be crumpled, rotated, or have handwritten notes. Which Azure AI service should you use to reliably extract this information?
80A company wants to use Azure AI Vision to extract text from scanned documents that contain both printed and handwritten text in multiple languages. The documents are large and can take several minutes to process. The solution must return the extracted text asynchronously. Which Azure AI Vision API should they use?
81A developer is using the Azure AI Vision Image Analysis API to extract text from a photo of a street sign. The sign contains text in both English and Japanese arranged in multiple columns. The developer needs the response to include the detected language and the bounding box for each text line. Which feature should they use?
82You are training an Azure Custom Vision object detection model to locate pallets in warehouse photos. Your training set contains 500 images, but only 40 images include pallets, while the rest are empty aisles. The model performs poorly, often missing pallets. You need to improve detection while keeping training time reasonable. What should you do first?
83You are developing a solution that uses Azure AI Vision to analyze images of products on an assembly line. You need to detect whether each product has a specific logo and also read the serial number printed on it. Which two Azure AI Vision services should you use? (Choose two.)
84A developer is building a mobile app that uses Azure AI Vision to generate a descriptive caption for user-uploaded photos. The app must return a human-readable sentence describing the main content of each image. Which Image Analysis feature should the developer use?
85A media company wants to automatically generate alt text for images on its news site using Azure AI Vision. The images are stored in Azure Blob Storage, and the solution must run serverless and respond within seconds. Which approach should you use?
86A company wants to build a mobile app that recognizes and tags landmarks in photos taken by tourists. They need a prebuilt model that requires no training and can identify thousands of famous places worldwide. Which Azure AI Vision feature should they use?
87You are building a mobile app that uses Azure AI Vision to generate captions for photos taken by users. The app must work offline in areas with no internet connectivity. Which Azure AI Vision feature should you use?
88You are using Azure AI Custom Vision to detect defects in fabric rolls. After training an object detection model, you notice that it often misses small tears. You have a large dataset of labeled images, but the tears are very small relative to the image size. What should you do to improve the model's detection of small tears?
89A developer needs to generate a descriptive caption for an image using Azure AI Vision. The image contains a dog catching a frisbee in a park. Which feature of Image Analysis should they use?
90A company wants to build a solution that automatically generates alt text for images on their website to improve accessibility. The alt text must be a concise, human-readable description of the image content. Which Azure AI Vision feature should they use?
Deep-dive questions
The most-searched questions in this domain — detailed explanations, worked examples, full answer breakdowns.
Be able to map a scenario to the right Azure Vision capability: Read for OCR, liveness for anti-spoofing, Custom Vision for classification/detection, and Docker export for edge deployment. Most important: match the feature to the exact requirement, not the most familiar service.
The Courseiva AI-102 question bank contains 90 questions in the Implement computer vision solutions domain, covering the 13% of the exam attributed to this domain in the official Microsoft blueprint. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the Implement computer vision solutions domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included