Skip to content
convertirdocs

Can't Copy Text From a PDF? Why It Happens and How to Fix It

If a PDF won't let you select or copy text, or pastes as gibberish, it's almost always a scan or a restriction. How to tell which one and fix it.

· 6 min read

You want to copy a clause from a contract, an account number from an invoice or a quote from a book, and the PDF won't cooperate: you drag the cursor and nothing highlights, or the whole page highlights as one block, or what you paste comes out as meaningless symbols. It's one of the most common PDF problems, and it almost always has one of four causes.

This guide helps you figure out which one you have in a minute, and fix it.

Quick diagnosis

Open the PDF and try this:

  1. Try to select a single word. If nothing highlights, or the whole page highlights like a rectangle, the PDF is probably an image (cause 1).
  2. If the word highlights but "Copy" is grayed out or does nothing, the document has copy restrictions (cause 2).
  3. If it copies but pastes as symbols, boxes or swapped letters, the problem is in how the PDF stores its fonts (cause 3).
  4. If it copies fine but comes out jumbled (columns mixed together, lines cut in half), the text is there but its internal order doesn't match the visual one (cause 4).

Cause 1: the PDF is an image (scanned)

This is the most common one. When a document is scanned or photographed, what ends up inside the PDF is a picture of each page. You see letters; the computer sees pixels, exactly like a photo of your dog. There's no text to copy.

It happens with documents scanned at an office, certificates someone printed and scanned again, phone photos converted to PDF, and many "copy" PDFs handed out by institutions.

The fix: OCR

OCR (optical character recognition) reads the image, identifies the letters and turns them into real text. With OCR PDF:

  1. Open the tool and choose your PDF.
  2. Tick the document's language: English, Spanish or Portuguese. If it mixes languages, tick each one. Picking the right language is what lets accented characters come out correctly.
  3. Run it. The first time, your browser downloads the language data (a few megabytes); your document isn't uploaded, recognition happens on your device.
  4. Download the result.

The PDF looks exactly as it did before, but now it has an invisible text layer over each page. You can search, select and copy.

If you want to edit, not just copy

When the goal is to change the content (fix a date, update a résumé you only have as a scan), run the OCR'd file through PDF to Word in editable mode. Without OCR first, a scan converted to Word comes out as images with no text.

If you only need plain text with no formatting, PDF to TXT pulls it out cleanly so you can paste it anywhere.

If you have a photo, not a PDF

For a single photo, a screenshot or an image someone sent you in a chat, there's no need to go through PDF: Image to text recognizes the text and puts it in an editable box for you to copy, or to download as .txt or Word.

What OCR can't do

  • Handwriting is recognized poorly. OCR is built for printed text.
  • A blurry, tilted or shadowed scan produces errors: confused letters, split words. If you can, rescan with Scan to PDF, cropping the corners and using a document filter.
  • Always check numbers, names and account numbers after copying them. An 8 read as a 3 in an account number is a real problem.

Cause 2: the PDF has copy restrictions

Some PDFs open without trouble but have an owner protection turned on that blocks copying text, printing or editing. It's common in bank statements, certificates, study guides and corporate documents. Your viewer will show Copy as disabled, or the text highlights but won't copy.

There are two kinds of protection to tell apart:

  • Restrictions (the PDF opens without asking for a password): these can be removed without typing any password. Remove PDF password gives you a copy without those limits, where copying works.
  • An open password (the PDF asks for a password before showing anything): here you do need to know it. Type it once in Unlock PDF and you get a copy that opens without it. The tool doesn't guess or crack passwords: if you don't know it, ask whoever sent you the file. For bank statements it's often a piece of your own data, and the bank usually says which in the email.

Only use this on documents you have the right to use this way: your own files or ones given to you for your use. Removing protection from someone else's document without permission can breach copyright or confidentiality agreements.

One detail: sometimes a PDF has both, restrictions and a scanned body. Unlock it first, then run OCR.

Cause 3: text copies as gibberish

You select, copy, paste, and get swapped letters, meaningless symbols or a row of little boxes. The PDF does have text, but the program that created it stored the characters with an internal encoding that doesn't say which letter each one is. On screen it looks fine because the viewer draws the shape of each letter; when you copy, you get the internal code.

The practical fix is to treat that PDF as if it were a scan: run OCR so it reads the page as an image and generates fresh text. In OCR PDF, untick "Skip pages that already have text": if you leave it ticked, the tool sees that text already exists (even if it's garbage) and does nothing.

After OCR, try copying again and check the result. If you just need a particular passage, you can also run Image to text on a screenshot of that part.

Cause 4: it copies, but comes out jumbled

The text is recognized correctly, but when you paste it, columns get mixed, lines break in the middle or headings land at the end. It happens with multi-column PDFs, brochures, academic papers and tables. A PDF doesn't store "paragraphs": it stores letters at positions, and the order they were written in isn't always the order you read them in.

What to try:

  • Copy in pieces: one column or paragraph at a time instead of the whole page.
  • PDF to Word in editable mode tries to rebuild paragraphs, headings and simple tables. With running text the result is usually very close to the original; with complex layouts you'll need some manual cleanup.
  • PDF to TXT if you only want the content to paste elsewhere and don't care about formatting.

Why do people send PDFs like this?

It's rarely on purpose. Many offices print, sign, stamp and scan because that's what their process requires. Others protect documents so they can't be altered. If you need the text often, it's worth asking for the original digital version: sometimes they have it and nobody offered.

All without uploading your documents

OCR, unlocking and conversions run inside your browser. The PDF isn't sent to any server, and neither is any password you type. That matters, because the things people most often need to copy from are contracts, bank statements and personal documents.

Frequently asked questions

Does OCR change how my PDF looks? No. It adds an invisible text layer; the page looks the same.

Does it work on a phone? Yes, in your phone's browser. Long documents take longer because recognition uses your phone's processor.

Which languages does it recognize? English, Spanish and Portuguese. You can tick several at once.

The PDF won't let me print or copy. Do I need the password? Not if the PDF opens without asking for one. Those restrictions are removed without typing anything.

Can I copy text from a phone photo? Yes, with Image to text. Take the photo straight on, in good light and close up so the letters look large.

Tools used in this guide