A developer needs to extract text from a PDF that contains only scanned images with no embedded text layer. The automation must run unattended on a machine without Microsoft Office installed. Which activity should the developer use to reliably obtain the text?
Read PDF with OCR renders each page as an image and applies an OCR engine to recognize text. This works for scanned PDFs without a text layer. UiPath supports multiple OCR engines that do not require Microsoft Office, so the activity can run unattended and reliably extract text from image-based PDFs.
Why this answer
Read PDF with OCR is specifically designed to handle PDFs that contain only images by rendering pages and applying OCR. It does not depend on Microsoft Office and can be configured with various OCR engines. This makes it the correct choice for unattended extraction from scanned PDFs without a text layer.
Exam trap
The trap here is assuming that Read PDF Text can handle any PDF, when it only reads embedded text and fails silently on scanned images.