The flight, hotel, and car confirmations came as three PDFs and the executive leaves tonight. They shouldn't have to o…
Open the scanned PDF file in Adobe Acrobat.
Select Scan & OCR from the right-hand pane.
Click Recognize Text and choose In This File.
Set the page range and language as needed, then click Recognize Text to apply OCR.
CautionEnsure OCR completes before proceeding; incomplete OCR may result in unsearchable or unselectable text in extracted pages.
After OCR finishes, select Organize Pages from the right-hand pane.
Choose Split to divide the document, or select the specific pages you want to extract.
If extracting, click Extract; if splitting, set your split criteria and confirm.
Save the new PDF(s) with a distinct file name.
Best practiceAfter extraction, review the output PDF to confirm that text is selectable and searchable, ensuring OCR was successful (BLOGS).
After OCR, always verify that the text in your split or extracted PDF is selectable to confirm successful recognition.
This runs in the Acrobat app - there is no separate API for this task.
To split a scanned PDF and extract the text, use the split tool to select the pages you want, then use OCR to turn the images into real text. This lets you save only the parts you need and copy or edit the text easily.
OCR stands for Optical Character Recognition. It is a tool that reads the picture of words in a scanned document and changes it into text you can select, copy, or edit on your computer.
Yes, you can choose specific pages to extract from a scanned document. After splitting out the pages you want, you can run OCR on just those pages to get the text.
In a normal PDF, the text is already selectable and editable. In a scanned document, the text is just an image, so you need OCR to turn it into real text before you can copy or edit it.
Using OCR lets you turn the images of text in your scanned document into real, usable text. This makes it much easier to search, copy, and edit the information you need from the split pages.