Create durable digital surrogates of fragile originals by converting the high‑resolution scans to searchable PDFs, exp…
Open the scanned PDF file in Adobe Acrobat.
Go to the Tools menu.
Select Scan & OCR from the tools list.
Click Recognize Text and choose In This File.
In the settings panel, ensure PDF Output Style is set to Searchable Image.
Click Recognize to start the OCR process.
Save the PDF after OCR is complete.
Best practiceFor best OCR accuracy, scan your documents at 300 dpi and ensure the images are clear before running OCR.
After OCR completes, you can search, copy, and highlight text in your previously scanned PDF.
This runs in the Acrobat app - there is no separate API for this task.
OCR, or Optical Character Recognition, is a technology that converts images of text into actual, editable text. It matters because when you scan a document, it's just a picture. OCR makes that picture's text readable by computers, so you can search for words, copy text, and even edit it. This is very important for archiving and indexing documents, making old paper files useful in a digital world.
Acrobat gives this error because the page you are trying to OCR already has real text, not just an image. The OCR feature is designed to work on pages that are only images. If the page already has text, Acrobat sees no need to add another text layer and stops the process. You only need OCR for image-only pages.
| Feature | Scanned PDF | OCR'd PDF |
|---|---|---|
| Content Type | Image of text | Image with hidden, searchable text |
| Searchable | No, you cannot search for words | Yes, you can search for words |
| Copy Text | No, you cannot copy text | Yes, you can copy text |
| File Size | Often larger (just an image) | Can be optimized, includes text data |