Make Scanned Forms Searchable with OCRmyPDF
Turn a scanned council, benefits, or school form into a searchable PDF on your own computer, so you can find the reference number in seconds.
Why this works
Government and school paperwork often arrives as scans. OCRmyPDF adds a text layer to the PDF using Tesseract, so you can search, copy, and paste from it without uploading anything.
Step by step
- Install OCRmyPDFFollow the project's install guide for your system. It needs Tesseract and Ghostscript, which the guide lists.
- Run it on one formUse
ocrmypdf --language eng scan.pdf searchable.pdf. Use--skip-textif some pages already have text. - Check the resultSearch for a word you know is on the form. If it is not found, rescan at higher resolution and repeat.
- Save the reference numbersCopy the reference, case number, or deadline into your calendar or a notes file the same day.
- Keep originalsStore the scan and the searchable copy together, with a date in the filename.
Copy-paste prompt
From the text of this official form, list: the issuing office, any reference or case number, every deadline with its date, the documents the form asks me to provide, and the contact details. Quote each item exactly. Do not interpret the legal effect of the form. FORM TEXT: <paste>
Check your result
- The reference number matches the printed form
- Every deadline is in your calendar
- The original scan is kept alongside the searchable copy
Pitfalls to avoid
- Poor scans give poor text; use 300 dpi or higher if you can.
- If a deadline matters, confirm it with the issuing office.
What we checked: Licence and project status were read from the project's GitHub repository on 8 Oct 2026. Install commands and flags change between releases, so follow the project README for the current version. Source: https://github.com/ocrmypdf/OCRmyPDF.