Convert the scanned traffic count report PDF into an Excel file and XML export so I can import the counts into my spre…
Open the scanned PDF or image file in Adobe Acrobat.
Go to the Tools menu and select Recognize Text.
In the Recognize Text dialog, click Recognize Text and choose In This File.
Adjust language and page settings if needed, then click Recognize Text to start OCR.
Best practiceFor best OCR accuracy, scan documents at 300 dpi and ensure the image is clear and not overly bright or dark.
After OCR completes, select and copy text directly from the PDF, or use Export PDF to save as an editable format such as Word.
NoteYou can now search, select, and edit the recognized text within Acrobat.
After OCR, you can highlight, copy, or edit the text just like a regular PDF document.
This runs in the Acrobat app - there is no separate API for this task.
An image-only page is like a photograph of a document; the text cannot be selected or searched. A page with renderable text already has a text layer, meaning the text can be highlighted, copied, and searched. Acrobat's OCR feature works by adding a searchable text layer to image-only pages, making the text usable. It does not need to run on pages that already have renderable text because the text is already there.
Acrobat shows this message because the page already has a layer of selectable text. OCR is for converting images of text into real text. If the text is already real, OCR is not needed. You can usually select and copy the text directly without running OCR again.
Acrobat's ability to extract tables from a blurry scan will be very low. OCR works best with clear, high-quality images. Blurry text is hard for the software to recognize, leading to many errors or no recognition at all. It is always better to rescan the document at a higher quality if possible to get accurate results.
After running OCR, Acrobat allows you to correct text recognition errors. You can go to 'All tools' > 'Scan & OCR' > 'Correct Recognized Text'. This tool highlights potential errors and lets you manually type in the correct words. It helps ensure the final text layer is accurate and matches the original document.