Courseiva

AI-102 Practice Question: Implement knowledge mining and information extraction solutions

You are configuring an Azure AI Search indexer to process documents from Azure Blob Storage. The documents include PDFs and Microsoft Word files. You need to extract both text and metadata such as author and creation date. Which indexer configuration should you use?

⚠ Common exam trap

The trap here is assuming that a specific parsing mode like text is needed for text extraction, when the default mode already handles common document formats with metadata.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

Set the parsingMode to default

The default parsingMode in Azure AI Search indexers is designed to handle a variety of document formats, including PDF and Microsoft Office files. It uses built-in document cracking to extract text and metadata, such as author and creation date, which are then available for mapping to index fields. Other parsing modes are specialized for JSON, delimited text, or plain text and do not support rich document extraction.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    Set the parsingMode to json

    Why it's wrong here

    The json parsingMode is used for JSON documents, where the indexer parses JSON structures. It is not suitable for PDFs or Word files, which are binary formats. Using this mode would result in no content extraction from those file types. Therefore, it does not meet the requirement to extract text and metadata from PDFs and Word files.

  • ✓

    Set the parsingMode to default

    Why this is correct

    The default parsingMode uses the built-in document cracking capabilities to extract text and metadata from various file formats, including PDF and Microsoft Office files. It automatically detects the file type and uses the appropriate extractor. This mode is designed for exactly this scenario, where you need to process multiple document types and extract both content and metadata fields like author and creation date.

  • ✗

    Set the parsingMode to delimitedText

    Why it's wrong here

    The delimitedText parsingMode is for plain text files with delimiters, such as CSV. It does not handle binary formats like PDF or Word. This mode would not extract text or metadata from those documents. It is intended for structured text files, not rich documents.

  • ✗

    Set the parsingMode to text

    Why it's wrong here

    The text parsingMode treats the entire file as plain text, ignoring any internal structure or metadata. It would not correctly extract text from binary formats like PDF or Word, and it would not retrieve metadata such as author or creation date. This mode is only suitable for simple text files.

Quick reference

Azure Blob Storage Tier Comparison

TierStorage CostRetrieval CostLatencyUse Case
HotHighestLowestImmediateActive data, frequent reads
CoolLowerHigherImmediateData accessed < once / month
ColdLower stillHigherImmediateData accessed < once / quarter
ArchiveLowestHighest + rehydration delayHoursLong-term compliance retention

About these practice questions

Courseiva writes every AI-102 question from scratch — 761 in total, each with an explanation and a wrong-answer breakdown. None are copied from real exams or dumps. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official Microsoft exam blueprint

This AI-102 practice question is part of Courseiva's free Microsoft certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the AI-102 exam.