I have a stack of songwriter annotations in pen on a lyric sheet that must be archived and made text-searchable for fu…
Open your scanned PDF document in Adobe Acrobat.
Go to the Tools menu in the top left menu bar.
Select Scan & OCR from the list of tools.
In the Scan & OCR pane, click Recognize Text.
Choose In This File to apply OCR to the open document.
Click Recognize Text to start the OCR process.
Save the PDF to retain the searchable text layer.
Best practiceFor best OCR accuracy, scan your documents at 300 dpi and avoid extreme brightness settings.
After OCR, you can search, select, and copy text from your scanned PDF as if it were a normal document.
This runs in the Acrobat app - there is no separate API for this task.
OCR, or Optical Character Recognition, is a technology that converts images of text into actual, editable text. It is important for scanned documents because most scans are just pictures. Without OCR, you cannot search for words, copy text, or edit the document. OCR makes your scanned documents useful for archiving and indexing.
If your scanned lyrics are blurry, the OCR accuracy will be lower. Acrobat needs clear text to recognize characters correctly. To improve results, try rescanning the lyrics at a higher resolution, like 300 dpi. Also, ensure good lighting when scanning and that the page is flat. If the blur is severe, manual correction might be needed after OCR.
Acrobat shows 'page contains renderable text' because the page already has real text, not just an image. The OCR feature is designed only for pages that are pure images. If your document has existing text, you do not need to run OCR on it.
| DPI Setting | OCR Accuracy | File Size |
|---|---|---|
| 72 dpi | Lower accuracy, especially for small or complex text. This is the minimum. | Smaller file size. |
| 300 dpi | Much higher accuracy, recommended for best results. | Larger file size. |