Scan and archive the signed release form returned by [client=Olivia Bennett] into her file today and save the complete…
Open the scanned PDF file in Adobe Acrobat.
Go to the Tools menu.
Select Scan & OCR from the tools list.
Click Recognize Text and choose In This File.
In the settings panel, ensure PDF Output Style is set to Searchable Image.
Click Recognize to start the OCR process.
Save the PDF after OCR is complete.
Best practiceFor best OCR accuracy, scan your documents at 300 dpi and ensure the images are clear before running OCR.
After OCR completes, you can search, copy, and highlight text in your previously scanned PDF.
This runs in the Acrobat app - there is no separate API for this task.
For the best OCR accuracy, you should scan your documents at 300 dpi. This resolution helps the software recognize characters more clearly. Scanning below 72 dpi will give poor results.
OCR is designed to work on image-only pages, meaning pages where the text is part of an image, not actual digital text. If a page already has 'renderable text,' it means the text is already digital and searchable. OCR cannot be applied to these pages because there's no image text to convert. It only processes visual text from scanned images.
Recognizing text in a scanned PDF creates a searchable layer. This means you can easily find specific words or phrases within the document. For archiving, it makes it simple to locate old documents quickly. For indexing, it allows systems to categorize and organize documents based on their content, improving retrieval efficiency.
| Feature | Regular Scanned PDF | OCR'd Scanned PDF |
|---|---|---|
| Text Selectable/Searchable | No, text is part of an image | Yes, text is selectable and searchable |
| Copy/Paste Text | Cannot copy text | Can copy and paste text |
| File Content | Image of a document | Image of a document with an invisible text layer |
| Accessibility | Limited for screen readers | Improved for screen readers |