Why OCR Rejects Your Document: 8 Common Reasons OCR Fails (And How to Fix Them)
August 2026
When OCR "rejects" a document, it's almost never the software refusing to process it — it's the document failing one of OCR's basic requirements. The same scan that fails in one tool usually fails in every tool, because the cause is in the file. Here are the eight most common reasons and how to fix each.
1. The scan is too blurry or low-resolution
If a human squinting at the page can't read it, OCR can't either. Fix: rescan at 300 DPI or higher, hold the phone steady, and make sure the document fills the frame.
2. It's a photo of a screen, not a scan
Photos of monitors pick up moiré patterns and glare that break text detection. Fix: use screenshots instead of photos of screens.
3. The page is skewed or rotated
Most modern tools auto-deskew, but extreme angles (90° sideways pages) can still fail. Fix: rotate the image to portrait before uploading, or photograph straight-on.
4. Poor contrast — faded ink, light pencil, dark backgrounds
Light gray text on white, or white text on dark backgrounds, confuses recognition. Fix: increase contrast in any photo editor, or scan with the "document" setting instead of "photo".
5. The document is encrypted or password-protected
Locked PDFs can't be read. Fix: unlock the PDF locally before uploading (a PDF "owner password" that blocks copying can usually be removed with free tools; an open password needs the password itself).
6. Text is behind stamps, watermarks or handwriting
"CONFIDENTIAL" stamps and redaction marks over text make the underlying characters unrecoverable. Fix: obtain a clean copy — no OCR can read what's covered.
7. Unusual fonts, tiny text, or dense tables
Decorative fonts, 6-point legal fine print and tightly packed tables stress recognition. Fix: zoom the page before scanning so text is a reasonable size; table-heavy pages often do better in tools with layout-aware OCR that exports to Excel.
8. The file is actually an image saved as "PDF"
This isn't a failure — the text simply isn't there yet. Image-only PDFs have no text layer, so copy-paste finds nothing. Fix: run the PDF through OCR — it reads the pixels and creates the text layer.
What to do when OCR fails
Try a better scan first — most failures are fixed at the source. If the document is fundamentally unreadable (faded thermal paper, heavy stamps), accept the limit and keep the original. If it's simply an image-PDF, OCR it with a layout-aware tool like CrazyOCR — that class of "failure" is actually the standard use case.
Frequently asked questions
Why does OCR fail on my scanned PDF?
The most common causes are low scan resolution (below ~300 DPI), blur, skewed pages, and poor contrast. Rescan with the document filling the frame in good light, and most failures disappear.
Why can't I copy text from a scanned PDF?
Scanned PDFs are images — they have no text layer. Copy-paste only works on text-based PDFs. Running the PDF through OCR creates the text layer and makes it selectable and searchable.
Can OCR read faded or low-quality documents?
Partly. OCR engines correct contrast and can rescue mildly faded prints, but if the ink has faded beyond what the eye can read, the information is gone — no software can recover it.
Why does OCR reject password-protected PDFs?
OCR software cannot open encrypted files. Remove the password locally first (with the password, or via PDF-unlock tools for owner-password files) before uploading.
