Be able to pick the right PDF activity for a given file and configure its page range. The most important thing is recognizing whether the PDF has a usable text layer: use Read PDF Text when it does, and OCR-based extraction when it does not.
Start practicing
PDF Automation — choose a session length
Free · No account required
Domain overview
This domain covers extracting data from PDFs in UiPath Studio using the Read PDF Text, Read PDF With OCR, and Digitize Document activities. Questions test choosing text versus OCR extraction, page-range properties, and how the Document Object Model behaves on scanned versus digitally created files.
Exam objectives
Selecting Read PDF Text versus Read PDF With OCR based on whether a PDF has a text layer
Using the Digitize Document activity with UiPath Document OCR to build a Document Object Model
Reading or extracting a labeled value such as an invoice number from a text-layer PDF
Setting the Range property in Read PDF Text to control which pages are read
Assuming Read PDF Text extracts text from scanned image-only PDFs, when those pages need OCR instead
Expecting Read PDF Text to return text embedded inside images on an otherwise digital PDF
Forgetting that the Range property accepts a page specification, so all pages are read by default
Click any question to see the full explanation and answer options, or start a focused practice session above.
When automating a PDF that contains both text-based data and scanned images, which activity approach ensures the highest accuracy for data extraction?
2Refer to the exhibit. What is the most likely cause of the error in the workflow?
3Which activity is most appropriate for extracting data from a PDF form with predefined fields?
4Which THREE factors can negatively impact the performance and accuracy of OCR-based PDF automation?
5Which property in the Read PDF Text activity is used to specify which pages to read?
6Refer to the exhibit. The activity executes without error but returns an empty string. What is the most likely reason?
7Which TWO of the following are valid ways to improve the accuracy of OCR in UiPath?
8You need to extract data from a PDF where the fields are located at different positions on every page. What is the most effective automation strategy?
9What is the primary function of the 'Read PDF Text' activity?
10What is the result of using the 'Read PDF Text' activity on a PDF that is both digitally created and contains embedded images?
11An automation project requires extracting specific invoice numbers from native PDF documents. Which approach ensures the most reliable performance while adhering to UiPath best practices for document processing?
12Which TWO of the following scenarios are best suited for using the 'Read PDF Text' activity instead of 'Read PDF with OCR'?
13When automating the extraction of data from a PDF that contains complex tables, which UiPath feature provides the most robust and structured results?
14A developer is configuring a workflow to extract all embedded raster images from a multi-page PDF contract for archiving purposes. Which activity is specifically designed to accomplish this task efficiently?
15A developer needs to extract text from a PDF that contains only scanned images with no embedded text layer. The automation must run unattended on a machine without Microsoft Office installed. Which activity should the developer use to reliably obtain the text?
16A developer uses the Digitize Document activity with the UiPath Document OCR engine on a scanned invoice PDF. The resulting Document Object Model returns correct text for printed fields, but several checkbox selections are reported as empty. Which action will most reliably capture the checkbox states?
17An invoice automation must extract line items from PDFs produced by three different vendors. Two vendors generate text-based PDFs, while the third sends scanned images. The developer needs one workflow that selects the correct extraction method per document without human intervention. Which approach meets this requirement?
18An automation must extract a specific invoice number from a PDF that has a text layer. The invoice number always appears after the literal label 'Invoice #:' on the first page. Which approach using UiPath PDF activities will most reliably isolate just the invoice number?
19A developer is configuring the Read PDF With OCR activity to process a batch of scanned PDFs that contain small fonts and low-contrast text. The goal is to maximize text recognition accuracy while keeping processing time reasonable. Which two settings should be adjusted to improve accuracy? (Choose two.)
20A developer needs to extract text from a PDF that contains a mix of digital text and scanned images of receipts. The digital text is machine-readable, and the scanned images contain only graphical content. Which UiPath approach will correctly extract text from both parts in a single automation?
21A developer needs to extract the text content of a digitally created PDF and store it in a string variable for later string manipulation. The PDF opens normally and its text can be selected in a viewer. Which activity should the developer use?
22A developer automates extraction of purchase order numbers from supplier PDFs. The Read PDF Text activity returns a long string containing the entire document. The developer needs only the value that immediately follows the label "PO Number:" on the same line. Which approach reliably isolates that value for downstream processing?
Be able to pick the right PDF activity for a given file and configure its page range. The most important thing is recognizing whether the PDF has a usable text layer: use Read PDF Text when it does, and OCR-based extraction when it does not.
The Courseiva UiPath-ADAv1 question bank contains 22 questions in the PDF Automation domain. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the PDF Automation domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included