How to make a searchable PDF
A PDF straight from a camera or basic scanner is a picture of text, not text. You cannot search it, copy from it, or select a word in it. OCR — optical character recognition — fixes that by reading the page and adding an invisible text layer. Here is what that means in practice and how to get it.
What you end up with: A PDF you can search by any word, copy text from, and select like a digital document.
- 1
Understand what OCR adds
OCR does not change how the document looks. It analyzes the page image, recognizes the characters, and stores them as a hidden text layer aligned behind the scan. The page still looks like paper; it just behaves like text.
- 2
Scan or import the document
In Paperly, OCR runs automatically when a page is scanned or imported — there is no separate 'recognize' step to remember. Both fresh scans and existing image-only PDFs can be processed.
- 3
Verify with a real search
Search for a word that appears once, deep in the document — a name, an invoice number, a clause keyword. If it is found, the text layer is working. This ten-second check beats trusting a progress bar.
- 4
Copy text where you need it
With OCR in place, text can be selected and copied straight from the PDF. Quoting a contract clause into an email stops being a retyping exercise.
- 5
Know the limits
Printed text recognizes reliably; clean handwriting often works; decorative fonts and heavy cursive may not. The scan is never harmed by a partial recognition — the image layer is always the original page.
Worth knowing
- OCR quality follows scan quality — a flat, shadow-free page recognizes far better than a curled, dim photo.
- On-device OCR means the text of medical forms, IDs, and financial statements is never sent to a recognition server.
- Searchable PDFs age well: an OCR'd archive is findable years later, an image-only one is a pile of pictures.
Do it in Paperly
This guide follows the real Paperly workflow. The relevant features, if you want the details: