Be able to design an Azure AI Search enrichment pipeline and pick the correct Document Intelligence model type. The critical skill is mapping skillset outputs into index fields and knowing when to retrain a custom model with more labeled documents.
Start practicing
Implement knowledge mining and document intelligence solutions — choose a session length
Free · No account required
Domain overview
This domain covers Azure AI Search pipelines (data sources, indexers, skillsets, indexes) and Azure AI Document Intelligence custom and prebuilt models. Questions test choosing the right cognitive skill, wiring enrichment to an index, and diagnosing extraction or confidence issues in knowledge mining solutions.
Exam objectives
Selecting built-in skillset skills such as Key Phrase Extraction, Language Detection, and OCR
Configuring indexers, data sources, and output field mappings from enriched documents
Choosing Document Intelligence prebuilt models (invoice, receipt, ID) versus custom models
Training and improving custom Document Intelligence models with labeled samples
Adding Key Phrase Extraction without Language Detection, so multilingual documents are not normalized before key phrase extraction runs.
Forgetting output field mappings or the knowledge store, so enriched skill outputs never reach the search index.
Assuming a custom Document Intelligence model generalizes; low-confidence or failing extractions usually mean more labeled samples are needed.
Click any question to see the full explanation and answer options, or start a focused practice session above.
A company is building a knowledge mining solution using Azure AI Search. They need to extract key phrases from a large set of documents in multiple languages. Which skill should they add to the skillset?
2A healthcare organization uses Azure Document Intelligence to process patient intake forms. They notice that the confidence scores for field extraction are low. What is the most likely cause?
3A company builds a knowledge mining solution using Azure AI Search with a custom skillset that includes an OCR skill. They want to ensure that images embedded in PDFs are processed. What should they configure?
4A law firm uses Azure Document Intelligence to extract clauses from legal contracts. They have a custom model trained on 15 labeled contracts. The model extracts clauses with high confidence on similar documents but fails to extract correct clauses from a new batch of contracts that have a different font and layout. The firm needs to improve extraction accuracy without retraining the model from scratch. The solution must minimize manual effort and cost. What should they do?
5Drag and drop the steps to set up Azure AI Content Safety for content moderation into the correct order.
6Match each Azure AI tool to its purpose.
7You are developing an Azure AI Search solution that indexes scanned PDF invoices. The indexer must extract text from the PDFs and also recognize entities such as organization names and dates. You want to use built-in cognitive skills to minimize custom code. Which combination of skills should you include in the skillset?
8A company uses Azure Document Intelligence with a custom neural model to extract data from purchase orders. The model was trained on 50 labeled samples and performs well on similar documents. However, when processing new purchase orders from a different supplier, the model fails to extract the 'TotalAmount' field accurately. What should you do to improve the model's performance on the new supplier's documents?
9You are building an Azure AI Search solution that uses a custom skill to enrich documents. The custom skill is implemented as an Azure Function. You need to ensure that the custom skill can access the documents' content and output enriched data. Which two actions should you perform? (Choose two.)
10You are building an Azure AI Search knowledge mining pipeline that enriches scanned PDF invoices. The PDFs are stored in Azure Blob Storage, and you need to extract text from each page before running downstream entity recognition. The solution must minimize development effort and rely on a built-in cognitive skill. Which skill should you add to the skillset?
11You are building a knowledge mining solution with Azure AI Search. The solution indexes scanned PDF reports stored in Azure Blob Storage. Each report contains multiple embedded images with text that must be searchable. You have already created a data source, index, and indexer. You need to ensure that text from the embedded images is extracted and mapped to the 'content' field in the index. Which two actions should you perform? (Choose two.)
12A company uses Azure Document Intelligence to process purchase orders. They have trained a custom model with 10 labeled samples and deployed it as 'purchaseOrderModel'. When analyzing a new purchase order, the extracted 'TotalAmount' field is often incorrect. The company wants to improve the model's accuracy for this field. What should they do?
13You are designing an Azure AI Search enrichment pipeline that processes scanned PDF invoices stored in Azure Blob Storage. You need to extract both printed text and handwritten notes from the documents before sending the content to an Azure AI Language entity recognition skill. The solution must minimize development effort and cost. Which skill should you add to the skillset?
14A company uses Azure AI Search to index documents from an Azure SQL Database. They have configured an indexer with a skillset that includes a custom skill hosted in an Azure Function. The custom skill enriches each document with a 'category' field. After running the indexer, they notice that the 'category' field is missing in the index for all documents. The Azure Function logs show that it is receiving requests and returning responses. What is the most likely cause?
15A logistics company uses an Azure AI Search indexer to process bills of lading stored in Azure Blob Storage. The indexer uses a skillset with a ShaperSkill that builds a complex object named 'shipment' containing nested fields for carrier, origin, and destination. After a full index run, queries for the carrier field return no results even though the source documents contain the data. You need to make the carrier value searchable. What should you do?
16A financial services firm needs to extract structured fields such as invoice date, vendor name, and total amount from thousands of PDF invoices. The documents vary in layout across vendors. The firm wants a pretrained model that requires no custom training and can return field-level confidence scores. Which Azure AI Document Intelligence model should they use?
17A company is building a knowledge mining solution using Azure AI Search. They need to extract text from handwritten notes stored as images in Azure Blob Storage. They want to use the built-in OCR skill in a skillset. Which cognitive service does the OCR skill rely on?
18You are building an Azure AI Search knowledge mining solution over a repository of scanned product manuals. You need to extract structured entities such as product names and part numbers from the OCR text and store them in an index field. (Choose two.)
19A media company is building a knowledge mining solution with Azure AI Search. They need to enrich video assets by extracting spoken words from the audio track and then indexing that transcript for search. Which two components must be included in the enrichment pipeline? (Choose two.)
20A company uses Azure AI Search to index a large collection of scanned invoices stored in Azure Blob Storage. They have a skillset that includes an OCR skill to extract text from the invoices. The indexer is configured to run every night. They notice that the indexer takes a long time to complete and sometimes times out. They want to optimize the indexer performance without reducing the quality of the extracted text. What should they do?
21A company uses Azure Document Intelligence to analyze prebuilt invoices. They need to extract the invoice total and the due date from each invoice. They call the Analyze Invoice operation and receive a result. Which part of the JSON response contains the extracted fields?
22A retail company wants to build a knowledge mining solution that indexes product descriptions stored in an Azure SQL Database and makes them searchable through a web application. The descriptions are already plain text. You need to configure Azure AI Search to pull the data into an index with the least effort. What should you create first?
23A company uses Azure Document Intelligence to process custom forms. They have trained a custom model using labeled data. They need to improve the model's accuracy for a specific field that is frequently misrecognized. The field appears in a consistent location but has varying formats. Which two actions should they take? (Choose two.)
24A company uses Azure AI Search to index documents from an Azure SQL Database. They need to ensure that deleted rows in the database are also removed from the search index during incremental indexing. They have configured the data source with change detection policies. What should they do to enable deletion detection?
25You are creating an Azure AI Search index for a knowledge mining solution. The index must support searching for documents by a required category field that can have one of five predefined values, and you want to enable faceted navigation on that field. Which index field configuration should you use?
26You are designing an Azure AI Search solution that uses an AI enrichment pipeline to extract text and key phrases from scanned documents. You need to ensure the pipeline can process image-heavy PDFs and produce searchable text and key phrases. Which two actions should you include in the skillset? (Choose two.)
Deep-dive questions
The most-searched questions in this domain — detailed explanations, worked examples, full answer breakdowns.
Be able to design an Azure AI Search enrichment pipeline and pick the correct Document Intelligence model type. The critical skill is mapping skillset outputs into index fields and knowing when to retrain a custom model with more labeled documents.
The Courseiva AI-102 question bank contains 26 questions in the Implement knowledge mining and document intelligence solutions domain, covering the 9% of the exam attributed to this domain in the official Microsoft blueprint. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the Implement knowledge mining and document intelligence solutions domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included