You are troubleshooting an Azure AI Search indexer that fails to index a PDF file stored in Azure Blob Storage. The error message indicates that the document is encrypted. What is the most likely cause and solution?
Trap 1: The indexer is not configured with the PDF parser; set the parsing…
PDF parsing is already handled by the blob indexer's built-in document extraction; parsing mode governs text versus layout extraction, not decryption. It tempts because parser misconfiguration does cause indexing failures, but the error explicitly names encryption, so the fix is supplying decryption credentials.
Trap 2: The file format is unsupported; convert to PDF/A
PDF is a supported format, and PDF/A conversion does not remove password protection, so the encrypted content still fails. It tempts because unsupported formats do break indexers, but the error names encryption; the actual fix is decrypting the file or configuring the indexer with an encryption key.
Trap 3: The file is too large; split it into smaller parts
File size triggers a different error about document limits, not encryption; splitting the PDF leaves the encryption intact. It tempts because oversized blobs genuinely fail indexing, but the stated cause is encryption, which requires removing the password or granting the indexer a decryption key.
- A
The indexer is not configured with the PDF parser; set the parsing mode
Why it fails: PDF parsing is already handled by the blob indexer's built-in document extraction; parsing mode governs text versus layout extraction, not decryption. It tempts because parser misconfiguration does cause indexing failures, but the error explicitly names encryption, so the fix is supplying decryption credentials.
- B
The file format is unsupported; convert to PDF/A
Why it fails: PDF is a supported format, and PDF/A conversion does not remove password protection, so the encrypted content still fails. It tempts because unsupported formats do break indexers, but the error names encryption; the actual fix is decrypting the file or configuring the indexer with an encryption key.
- C
The file is too large; split it into smaller parts
Why it fails: File size triggers a different error about document limits, not encryption; splitting the PDF leaves the encryption intact. It tempts because oversized blobs genuinely fail indexing, but the stated cause is encryption, which requires removing the password or granting the indexer a decryption key.
- D
The PDF is encrypted; remove encryption before indexing
Azure AI Search's document cracking cannot decrypt password-protected or rights-managed PDFs, so the indexer reports the document as encrypted and skips it. Removing encryption before indexing, or supplying an unencrypted copy, lets the blob indexer extract text successfully.