Troubleshooting
How to Tell If a PDF Is a Scanned Document
Quick tests to identify a scanned PDF and one-click OCR to make it searchable. Free, in-browser, no upload, no sign-up.
Hamza M.
PDF Workflow Editor
A scanned PDF is a picture of a page. The letters on it are pixels, not characters, which is why you cannot select a word, copy a sentence or search with Ctrl+F — there is simply no text stored in the file. Two PDFs that look identical on screen can behave completely differently, and knowing which kind you are holding decides what you can do next.
The three ten-second tests
- Try to select a word. In a digital PDF, dragging across text highlights individual words and letters. In a scan, the whole page selects like one big picture.
- Search with Ctrl+F or Cmd+F for a distinctive word you can see. A real text layer finds it instantly; a picture-only page returns nothing.
- Zoom to 300–400 percent. Vector text stays crisp as it zooms, while a scan pixelates like a photograph. This is a useful hint, not proof — building a searchable OCR layer under a scan is exactly why the picture can still look sharp.
One test alone can lie. Searchable scans — an image with an invisible OCR text layer behind it — pass the selection and search tests while still being pictures at heart. The reliable read is the combination: no selectable words, no search results and visible pixelation together mean a scan.
What a scan means for editing and copying
If a PDF is a scan, text tools around it fail for the same reason: there are no characters to read. Copying produces nothing, translation returns the file untouched or empty, and an editor cannot touch the letters. The fix in every case is the same — run OCR to add a real text layer.
Make a scanned PDF searchable with OCR
- 1Open the OCR tool and drop in your scanned PDF.
- 2Wait while each page is recognised into a hidden text layer.
- 3Download the result — it looks identical, but the words are now real.
After OCR the same tests pass: selection works, Ctrl+F finds words, and copy, translation and editing behave like a digital document. The page image itself is unchanged.
When a PDF is mixed
Reports and applications often combine digital cover pages with scanned receipts, signed forms or appendices. The tests above give a per-page verdict, so a file that searches fine on page 1 and fails on page 5 is mixed — run OCR and only the image pages are affected, the digital text stays untouched.
Do it now
OCR PDF
Extract text from scanned documents. It runs in this browser tab — your file is not uploaded anywhere.
Open OCR PDFFrequently asked questions
How do I know if a PDF is a scan or real text?+
Try to select a single word and search with Ctrl+F. If you cannot highlight individual words and a word you can see is not found, the page is a picture. Two tests together give a reliable answer.
Can a scanned PDF pass the text test anyway?+
Yes. Scans that already have an invisible OCR layer behave like digital documents — and if that layer is poor, copied text comes out garbled. Re-running OCR fixes the layer.
What can I do with a scanned PDF once I know?+
Run OCR to add a text layer. The page looks the same but suddenly selection, search, copy, translation and editing all work.
Are some pages of my PDF scans and others not?+
Mixed files are common — a digital report with scanned receipts tacked on. Test a few pages, including the attachments area, to map which pages are pictures.
Does OCR change how the scan looks?+
No. OCR adds an invisible text layer underneath the page image, so the document still looks exactly like the original scan.