Why Can't I Copy Text From a PDF? 3 Causes and Their Fixes

Copy and paste failing in a PDF? It's one of three things: a scanned image, copy restrictions, or a broken font map. How to tell which, and the fix for each.

Lewis Hadden3 min read

Copy and paste is the least glamorous feature a document can support, which makes it all the more annoying when a PDF refuses. You drag across a paragraph and nothing selects - or everything selects, and pastes as □□□▯▯. There are exactly three common causes, each with a different fix, and you can diagnose yours in about ten seconds.

The ten-second diagnosis

Try to select a sentence, then try Ctrl+F for a word you can see:

What happensCause
Nothing selects, search finds nothingScanned image - no text layer
Text selects, paste comes out as gibberishBroken font mapping
Selection or copy is blocked outrightCopy restrictions

Cause 1 - The PDF is a scanned image

The most common case by far. A PDF made by a scanner, a phone camera, or a fax machine contains pictures of pages, not characters. There is literally nothing to copy - the "text" you see is pixels.

The fix is OCR, which recognises the characters in the image and adds them as a real text layer. Your options, in rough order of effort:

Cause 2 - The text copies, but pastes as gibberish

Sometimes selection works fine and the paste is garbage: wrong letters, symbol soup, or empty boxes. This is a font problem, not a scanning problem. PDFs store text as internal character codes plus instructions for drawing them; a separate table (ToUnicode) maps those codes back to real characters for copy and search. When a PDF generator embeds a font without that table - common with subsetted fonts, old design tools, and some print drivers - the page renders perfectly but copies nonsense.

The workaround is to stop trusting the broken text layer and read the rendered page instead: OCR the document as if it were a scan. An OCR pass reads what the page looks like, which is exactly the thing that's still correct. Any of the routes from cause 1 work here too.

Cause 3 - The PDF restricts copying

PDFs can carry a permissions flag that disallows copying while still letting you open and read the file. Some viewers enforce it strictly, others don't, which is why the same file behaves differently in different apps.

Two honest points about this case:

  • The restriction was set deliberately. If it's a colleague's or vendor's document, the right move is asking for an unrestricted copy - there's usually no drama, the flag is often a default someone never chose.
  • Restricted is not the same as encrypted. A file you can't open without a password is a different situation from one you can read but not copy from.

When what you actually want is the answer, not the clipboard

A lot of copy-paste frustration is a means to an end: you were going to paste the text into a chatbot, a summary, or an email to ask about it. If that's the real goal, you can skip the clipboard entirely. Sidenote reads the PDF you have open in the browser - scanned, broken-font, or born-digital - and answers questions about it directly, with every claim carrying a citation that scrolls to the exact passage. The text you couldn't copy becomes text you can quote, with the source one click away.

Frequently asked questions
The PDF's font is missing the mapping table (ToUnicode) that translates its internal character codes back to real letters. The page renders fine because the shapes are intact, but copy produces gibberish because the translation is broken. OCR-based extraction reads the rendered page instead, which sidesteps the broken mapping.
If it's permission-flagged, the honest first step is asking the document owner for an unrestricted copy - the restriction was set on purpose. Note that copy-restricted is different from password-locked: a PDF you can open but not copy from is enforcing a permissions flag that some viewers respect and others ignore.
No. A scanned page stores no characters at all, only an image, so there is nothing to copy until OCR recognises the text. Any tool that appears to copy from a scan is running OCR somewhere along the way.
For a born-digital PDF, a free extractor like our PDF to text tool pulls the embedded text in one step. For a scanned or broken-font PDF, use an OCR-based reader - Sidenote OCRs the page on ingest and lets you copy, quote, and question the text with citations.
All guides
Ready when you are

Stop digging. Start asking.

Add Sidenote to your browser, open any page in your wiki, and ask it the question you’ve been Slacking the team about.

7-day Pro trial · No card required · Free plan forever