Courseiva

AI-102 · domain

Implement knowledge mining and document intelligence solutions

This domain covers Azure AI Search pipelines (data sources, indexers, skillsets, indexes) and Azure AI Document Intelligence custom and prebuilt models. Questions test choosing the right cognitive skill, wiring enrichment to an index, and diagnosing extraction or confidence issues in knowledge mining solutions.

26 questions6 easy15 medium5 hard

Focused practice

Practice Implement knowledge mining and document intelligence solutions questions

Scored sessions drawing only from this domain — pick a length below.

What this domain covers

What to know about Implement knowledge mining and document intelligence solutions

Be able to design an Azure AI Search enrichment pipeline and pick the correct Document Intelligence model type. The critical skill is mapping skillset outputs into index fields and knowing when to retrain a custom model with more labeled documents.

Selecting built-in skillset skills such as Key Phrase Extraction, Language Detection, and OCR

Configuring indexers, data sources, and output field mappings from enriched documents

Choosing Document Intelligence prebuilt models (invoice, receipt, ID) versus custom models

Training and improving custom Document Intelligence models with labeled samples

Watch out for

Common Implement knowledge mining and document intelligence solutions exam traps

  • ▸Adding Key Phrase Extraction without Language Detection, so multilingual documents are not normalized before key phrase extraction runs.
  • ▸Forgetting output field mappings or the knowledge store, so enriched skill outputs never reach the search index.
  • ▸Assuming a custom Document Intelligence model generalizes; low-confidence or failing extractions usually mean more labeled samples are needed.

Question index

All Implement knowledge mining and document intelligence solutions questions (26)

Click any question to see the full explanation, or start a practice session above.

1

A financial services firm needs to extract structured fields such as invoice date, vendor name, and total amount from thousands of PDF invoices. The documents vary in layout across vendors. The firm wants a pretrained model that requires no custom training and can return field-level confidence scores. Which Azure AI Document Intelligence model should they use?

Easy
2

A company uses Azure Document Intelligence to process purchase orders. They have trained a custom model with 10 labeled samples and deployed it as 'purchaseOrderModel'. When analyzing a new purchase order, the extracted 'TotalAmount' field is often incorrect. The company wants to improve the model's accuracy for this field. What should they do?

Medium
3

Drag and drop the steps to set up Azure AI Content Safety for content moderation into the correct order.

Medium
4

You are designing an Azure AI Search enrichment pipeline that processes scanned PDF invoices stored in Azure Blob Storage. You need to extract both printed text and handwritten notes from the documents before sending the content to an Azure AI Language entity recognition skill. The solution must minimize development effort and cost. Which skill should you add to the skillset?

Medium
5

A logistics company uses an Azure AI Search indexer to process bills of lading stored in Azure Blob Storage. The indexer uses a skillset with a ShaperSkill that builds a complex object named 'shipment' containing nested fields for carrier, origin, and destination. After a full index run, queries for the carrier field return no results even though the source documents contain the data. You need to make the carrier value searchable. What should you do?

Hard
6

A company uses Azure AI Search to index documents from an Azure SQL Database. They have configured an indexer with a skillset that includes a custom skill hosted in an Azure Function. The custom skill enriches each document with a 'category' field. After running the indexer, they notice that the 'category' field is missing in the index for all documents. The Azure Function logs show that it is receiving requests and returning responses. What is the most likely cause?

Medium
7

You are creating an Azure AI Search index for a knowledge mining solution. The index must support searching for documents by a required category field that can have one of five predefined values, and you want to enable faceted navigation on that field. Which index field configuration should you use?

Easy
8

A law firm uses Azure Document Intelligence to extract clauses from legal contracts. They have a custom model trained on 15 labeled contracts. The model extracts clauses with high confidence on similar documents but fails to extract correct clauses from a new batch of contracts that have a different font and layout. The firm needs to improve extraction accuracy without retraining the model from scratch. The solution must minimize manual effort and cost. What should they do?

Medium
9

A company is building a knowledge mining solution using Azure AI Search. They need to extract key phrases from a large set of documents in multiple languages. Which skill should they add to the skillset?

Medium
10

A company uses Azure Document Intelligence to process custom forms. They have trained a custom model using labeled data. They need to improve the model's accuracy for a specific field that is frequently misrecognized. The field appears in a consistent location but has varying formats. Which two actions should they take? (Choose two.)

Medium
11

Match each Azure AI tool to its purpose.

Medium
12

You are designing an Azure AI Search solution that uses an AI enrichment pipeline to extract text and key phrases from scanned documents. You need to ensure the pipeline can process image-heavy PDFs and produce searchable text and key phrases. Which two actions should you include in the skillset? (Choose two.)

Medium
13

A company uses Azure Document Intelligence to analyze prebuilt invoices. They need to extract the invoice total and the due date from each invoice. They call the Analyze Invoice operation and receive a result. Which part of the JSON response contains the extracted fields?

Easy
14

A company uses Azure Document Intelligence with a custom neural model to extract data from purchase orders. The model was trained on 50 labeled samples and performs well on similar documents. However, when processing new purchase orders from a different supplier, the model fails to extract the 'TotalAmount' field accurately. What should you do to improve the model's performance on the new supplier's documents?

Hard
15

You are building an Azure AI Search solution that uses a custom skill to enrich documents. The custom skill is implemented as an Azure Function. You need to ensure that the custom skill can access the documents' content and output enriched data. Which two actions should you perform? (Choose two.)

Medium
16

A company uses Azure AI Search to index a large collection of scanned invoices stored in Azure Blob Storage. They have a skillset that includes an OCR skill to extract text from the invoices. The indexer is configured to run every night. They notice that the indexer takes a long time to complete and sometimes times out. They want to optimize the indexer performance without reducing the quality of the extracted text. What should they do?

Hard
17

A healthcare organization uses Azure Document Intelligence to process patient intake forms. They notice that the confidence scores for field extraction are low. What is the most likely cause?

Easy
18

A company is building a knowledge mining solution using Azure AI Search. They need to extract text from handwritten notes stored as images in Azure Blob Storage. They want to use the built-in OCR skill in a skillset. Which cognitive service does the OCR skill rely on?

Easy
19

You are building a knowledge mining solution with Azure AI Search. The solution indexes scanned PDF reports stored in Azure Blob Storage. Each report contains multiple embedded images with text that must be searchable. You have already created a data source, index, and indexer. You need to ensure that text from the embedded images is extracted and mapped to the 'content' field in the index. Which two actions should you perform? (Choose two.)

Hard
20

A company builds a knowledge mining solution using Azure AI Search with a custom skillset that includes an OCR skill. They want to ensure that images embedded in PDFs are processed. What should they configure?

Medium
21

You are building an Azure AI Search knowledge mining pipeline that enriches scanned PDF invoices. The PDFs are stored in Azure Blob Storage, and you need to extract text from each page before running downstream entity recognition. The solution must minimize development effort and rely on a built-in cognitive skill. Which skill should you add to the skillset?

Medium
22

A media company is building a knowledge mining solution with Azure AI Search. They need to enrich video assets by extracting spoken words from the audio track and then indexing that transcript for search. Which two components must be included in the enrichment pipeline? (Choose two.)

Medium
23

A company uses Azure AI Search to index documents from an Azure SQL Database. They need to ensure that deleted rows in the database are also removed from the search index during incremental indexing. They have configured the data source with change detection policies. What should they do to enable deletion detection?

Hard
24

You are building an Azure AI Search knowledge mining solution over a repository of scanned product manuals. You need to extract structured entities such as product names and part numbers from the OCR text and store them in an index field. (Choose two.)

Medium
25

You are developing an Azure AI Search solution that indexes scanned PDF invoices. The indexer must extract text from the PDFs and also recognize entities such as organization names and dates. You want to use built-in cognitive skills to minimize custom code. Which combination of skills should you include in the skillset?

Medium
26

A retail company wants to build a knowledge mining solution that indexes product descriptions stored in an Azure SQL Database and makes them searchable through a web application. The descriptions are already plain text. You need to configure Azure AI Search to pull the data into an index with the least effort. What should you create first?

Easy

Frequently asked questions

What does the Implement knowledge mining and document intelligence solutions domain cover on the AI-102 exam?
Be able to design an Azure AI Search enrichment pipeline and pick the correct Document Intelligence model type. The critical skill is mapping skillset outputs into index fields and knowing when to retrain a custom model with more labeled documents.
How many questions are in this domain?
This page lists all 26 Implement knowledge mining and document intelligence solutions questions in the AI-102 question bank. The actual exam draws from this domain proportionally to its weighting in the official exam blueprint.
What is the best way to practise this domain?
Start with a short focused session (10 questions) to identify gaps, then work through explanations. Repeat with a longer session once the weak areas feel solid.
Can I practise only Implement knowledge mining and document intelligence solutions questions?
Yes — the session launcher on this page filters questions to this domain only. Choose any session length for inline explanations and scoring.
ai-102 AI-102 knowledge mining doc intelligence Practice Questions