Run OCR on [document=ScannedFieldForms_Apr2025.pdf] so text is selectable and searchable, then save a searchable PDF […
Open the scanned PDF file in Adobe Acrobat.
Go to the Tools menu.
Select Scan & OCR from the tools list.
Click Recognize Text and choose In This File.
In the settings panel, ensure PDF Output Style is set to Searchable Image.
Click Recognize to start the OCR process.
Save the PDF after OCR is complete.
Best practiceFor best OCR accuracy, scan your documents at 300 dpi and ensure the images are clear before running OCR.
After OCR completes, you can search, copy, and highlight text in your previously scanned PDF.
This runs in the Acrobat app - there is no separate API for this task.
A searchable text layer is invisible text placed over the image of your scanned document. It matters because it allows you to use the search function to find words and phrases within the document, just like a regular digital file. Without this layer, your scanned form is just a picture, and you cannot search its content.
Acrobat's 'Recognize Text' feature (OCR) is designed to work only on pages that are images, like those from a scanner. If a page already has real, selectable text, Acrobat will not apply OCR to it. The error 'page contains renderable text' means the program thinks the page already has text, so it stops the process. This prevents it from trying to 'recognize' text that is already there.
| Feature | Regular PDF | Scanned PDF (after OCR) |
|---|---|---|
| Text Selectable | Yes, text is native and selectable. | Yes, text is selectable via an added invisible layer. |
| Searchable | Yes, all text is searchable. | Yes, all recognized text is searchable. |
| Origin | Created digitally (e.g., from Word, web page). | Created from an image (e.g., scanned paper document). |
| File Structure | Contains text characters directly. | Contains an image with a text layer on top. |