Convert scanned discovery and signed client letters for case [case=14-CV-310] into searchable text, run OCR on each PD…
Open the scanned PDF file in Adobe Acrobat.
Go to the Tools menu.
Select Scan & OCR from the tools list.
Click Recognize Text and choose In This File.
In the settings panel, ensure PDF Output Style is set to Searchable Image.
Click Recognize to start the OCR process.
Save the PDF after OCR is complete.
Best practiceFor best OCR accuracy, scan your documents at 300 dpi and ensure the images are clear before running OCR.
After OCR completes, you can search, copy, and highlight text in your previously scanned PDF.
This runs in the Acrobat app - there is no separate API for this task.
This message means the page already has actual text data, not just an image of text. OCR is designed for image-only pages. If a page already has text, Acrobat doesn't need to recognize it because it's already there and searchable. You should only use OCR on documents that are purely images.
| Feature | Regular PDF (image-only) | OCR'd PDF |
|---|---|---|
| Text Search | Cannot search for words | Can search for words |
| Text Select/Copy | Cannot select or copy text | Can select and copy text |
| File Type | Often a scanned image | Image with a hidden text layer |
Choosing the correct language is very important because OCR software uses language-specific dictionaries and rules to recognize characters. If you select the wrong language, the OCR engine might misinterpret letters, especially those with accents or unique shapes, leading to many errors and poor text recognition. This makes the document less searchable and less accurate.