Why Can’t I Select Text in a PDF?
If you cannot select text in a PDF — clicking and dragging over a word highlights nothing — the text is not in the file's content stream. That usually means one of four things: the PDF is a scanned image, the text was baked into an image when the file was exported, the file was flattened in a way that lost the text layer, or the font encoding is broken. The fix for the first three is OCR. This guide explains how to tell which cause applies and what to do.
The 4 common causes
1. The PDF is scanned (image-only)
The most common cause. A scanned document is a photograph of paper — each page is one big image. There is no text in the content stream, so nothing to select. Scanners produce image-only PDFs by default; the "Searchable PDF" option in scanner software is the one that adds a text layer via OCR at scan time.
How to check: zoom to 400%. If the text becomes pixelated (you see jagged edges around letter shapes), it is an image. Real text stays crisp at any zoom because it is rendered from font outlines. Also try Ctrl+F / Cmd+F — if Find cannot locate a word you can clearly see, the file has no text layer.
Fix: run OCR to add a searchable text layer. SeeHow to Make a PDF Searchable for the tools that do this locally.
2. Text was baked into an image
Some applications export PDFs where text is rasterized — converted to pixels and embedded as an image, even though the source was real text. This happens with "Print to PDF" drivers that rasterize everything, with presentation exporters that flatten text onto background images, and with PDFs that have been through a "flatten" step that converts all content to images.
How to check: same as scanned — zoom to 400% and look for pixelation. The file may have been a real text PDF originally; the rasterization happened during export.
Fix: re-export from the source application with text preserved as text (not rasterized). If you do not have the source, run OCR on the existing file.
3. The PDF was flattened
Flattening converts form fields, annotations, and sometimes text into static content. Some flatten operations preserve the text layer; others rasterize everything. If the flatten step rasterized, the result is image-only and non-selectable.
How to check: zoom test again. If the text is still crisp at 400% but you cannot select it, the flatten preserved text as outlines (vector shapes) rather than characters — rare but possible. If it is pixelated, the flatten rasterized.
Fix: re-flatten with a tool that preserves the text layer, or re-export from the source. SeeHow to Flatten a PDFfor the difference between text-preserving and rasterizing flatten.
4. Broken font encoding (ToUnicode missing)
The rarest cause. The text exists in the content stream and renders correctly, but the font's ToUnicode mapping is missing or wrong. Viewers can display the glyphs (they have the font outlines) but cannot map them back to Unicode characters for selection, copy, or search. You see the text but cannot select it.
How to check: the text looks sharp at any zoom (it is real text, not an image) but you cannot select or copy it. Find also fails. This is common with PDFs generated by old CAD tools or niche exporters.
Fix: re-export from the source with proper Unicode mapping. If you do not have the source, run OCR — it will add a fresh text layer with correct Unicode, replacing the broken one.
Quick diagnostic checklist
- Try to select a word. Click and drag over a word. If nothing highlights, the text is not in the content stream.
- Zoom to 400%. If the text becomes pixelated, it is an image (scanned, rasterized, or flattened to image). If it stays crisp, it is real text with a broken encoding.
- Try Find (Ctrl+F / Cmd+F). If Find cannot locate a word you can see, the file has no searchable text layer.
- Check file size. Scanned PDFs are usually large (one full-page image per page). A small file with non-selectable text is more likely a broken-encoding case.
- Apply OCR if needed. Use a local OCR tool to add a searchable text layer.
A quick test you can run right now
If your PDF was created from a scanned document, text extraction will return empty. Try IXPDF's PDF to Text tool — drop your file in, run it, and inspect the result. If the extracted text is empty, your PDF is scanned (or rasterized) and needs OCR (optical character recognition) to make the text selectable; see How to Make a PDF Searchable for trusted local OCR options. If the tool does return text but you still cannot select it in the original PDF, the cause is likely broken font encoding (cause 4 above) — the text layer exists, but the viewer cannot map glyphs back to Unicode.
Why OCR is the fix (and not just a workaround)
For scanned, rasterized, and broken-encoding PDFs, OCR is the only fix — there is no original text to recover. OCR analyzes the image, recognizes letter shapes, and writes a new text layer into the PDF. The result is a "searchable PDF": the original image stays visible (so the layout and signatures are preserved), and the recognized text sits behind it as an invisible layer for selection, copy, and search.
The quality of OCR determines the quality of the result. Modern OCR engines (Tesseract 5, ABBYY FineReader, Adobe Acrobat's OCR) are good on clean 300-DPI scans of common fonts. They struggle on handwriting, low-resolution scans, decorative fonts, and multi- column layouts with mixed text and images. Always proofread the recognized text before relying on it.
Common mistakes
- Assuming the PDF is "encrypted" or "protected."Non-selectable text is almost never a security feature. It is usually a scanned file. Check by opening in Acrobat Reader — if there are no restrictions listed in File > Properties > Security, the file is not protected.
- Using a free online OCR service for a sensitive document. The file is uploaded to a server. Use a local tool — Tesseract, Acrobat, or macOS Live Text.
- Trusting OCR output without proofreading. OCR misreads. If you are copying text into a contract or a citation, verify every word against the image.
- Re-scanning instead of re-exporting. If you have the source document (Word, PowerPoint, InDesign), re-export as PDF — the text will be real and selectable. Re-scanning is a last resort.
Make your PDF searchable
IXPDF cannot do OCR in the browser today — accuracy is not good enough to ship. See our guide for trusted local OCR tools that run on your machine without uploading your file.
Learn about OCR optionsTry it now
Ready to put this into practice?
Run the tool locally in your browser — no upload, no account, no watermarks. Your file never leaves your device.