Skip to content
Start scanning

How to Copy Text From a Scanned PDF

You cannot copy text from scanned PDF pages directly because each page is a photograph, not text. The fix is OCR: it reads the image and gives you real characters that you can copy and paste. Open the file in PDFNova OCR PDF, choose the document language, run recognition and use the copy button to put the text on your clipboard.

The whole thing takes a few minutes and happens on your own device. Below you will find the exact steps, what to do when you only need one paragraph, and how to tidy the text after pasting.

Why select and copy do nothing in a scan

When a document is scanned or photographed, the PDF stores an image of the paper. Your PDF reader sees colored dots, not letters, so there is nothing for it to highlight. Dragging across a line either selects nothing or draws a box around the whole page.

Some PDFs also have copying disabled by their owner, which is a different situation. If you can highlight individual words but the copy command is blocked, the file has permission settings and is not a scan. This guide is about scans.

Check whether you need OCR at all

What you see What it means What to use
Words highlight one by one Digital PDF with real text PDF to Text, or copy in your reader
Nothing highlights, or the whole page does Scanned PDF OCR PDF
Some pages highlight, others do not Mixed document OCR for the scanned pages
You have a photo or screenshot, not a PDF Image file OCR works on images too

For the last case there is a dedicated walkthrough in how to extract text from an image.

How to copy text from a scanned PDF step by step

  1. Open the OCR PDF tool in your browser on a phone, tablet or computer.
  2. Add your scanned PDF. A password-protected file will ask for its password when it opens.
  3. Pick the language of the text: English, Urdu, Arabic, Hindi, French, Spanish, German or Chinese (Simplified).
  4. Start recognition. On first use of a language, its data downloads once, so allow a little extra time.
  5. When the text appears, use copy text to send it to your clipboard, or download it as a .txt file if you want to keep all of it.
  6. Paste into your document, email or notes app and proofread it against the original page.

If you also want to keep the document itself in a usable form, download the searchable PDF as well. After that you can open it in any PDF reader and copy from it normally, whenever you like.

When you only need one page or one paragraph

Running recognition on a 200-page scan to copy three lines wastes time, especially on a phone. Cut the job down first. Use Split PDF to pull out just the page or range you need, for example page 14 or pages 14-16, and run OCR on that small file.

From the recognised text, select and copy only the passage you want. Processing a small part is faster and makes it easier to compare the text with the page.

Clean up the pasted text

Recognised text is rarely perfect. A quick clean-up makes it ready to use.

  • Line breaks. OCR often ends a line where the printed line ended. Join the lines of a paragraph so the text flows properly.
  • Hyphenated words. A word split across two lines in print may arrive as two pieces with a hyphen. Rejoin them.
  • Look-alike characters. Watch for the letter O and zero, lowercase l and the number 1, and the pair rn read as m.
  • Numbers and names. Check every amount, date, phone number and surname against the image. These are the errors that cause real trouble.
  • Headers and page numbers. Remove running headers, footers and page numbers that got mixed into the paragraph.
  • Columns and tables. Text from two columns can be merged line by line. Copy each column separately if the order looks wrong.

Paste as plain text when your editor offers the choice. It avoids odd fonts and spacing and lets your own document style apply.

Copying Urdu, Arabic and other right-to-left text

Right-to-left scripts add two challenges. The first is recognition: connected scripts are harder than separate printed letters, and Urdu Nastaliq is more difficult than printed English, so expect more corrections. The second is display: correctly recognised text can look reversed or jumbled if the program you paste into is set to left-to-right.

Set the paragraph direction to right-to-left in your editor and use a font that supports the script. More detailed help is in our guide to working with right-to-left PDFs in Urdu and Arabic. If you need to turn the corrected text back into a document, Text and HTML to PDF supports right-to-left text.

If the copied text is wrong or empty

  • Nonsense characters. The language setting does not match the document. Change it and run again.
  • Many small mistakes. The scan is faint, blurry or tilted. Our list of tips on how to improve OCR accuracy shows what to fix before you retry.
  • Very little text found. The page may be sideways or upside down, or it may be handwriting, which OCR reads poorly.
  • Parts of the page missing. Text over photos, stamps or colored backgrounds is often skipped. Type those parts by hand.
  • It takes a long time. Recognition runs on your device. Process fewer pages at a time and keep the tab open.

Use copied text responsibly

Being able to copy text does not change who owns it. Quote with credit, follow your school or employer’s rules on reuse, and remember that copyright rules differ by country. Also take care with personal data in scanned forms and letters. PDFNova processes files in your browser, so your documents are not uploaded to a server, but what you paste and share afterwards is still your responsibility.

Frequently asked questions

Why can I not copy text from my scanned PDF?

Because the pages are images. There is no text in the file until OCR recognises it.

How do I copy text from a scanned PDF on a phone?

Open the OCR tool in your mobile browser, add the PDF, choose the language, run recognition and tap copy. Then paste into any app.

Will the copied text keep its formatting?

No. You get plain text. Bold, fonts, columns and tables are not preserved, so apply formatting again where you paste it.

Can I copy text from a scanned PDF without retyping?

Yes for printed text in a clear scan. You will still need to proofread, and handwritten pages usually need retyping.

Next time a scan refuses to let you highlight a line, open OCR PDF, run it on the pages you need and paste the text where you want it.

Leave a Comment

Your email address will not be published. Required fields are marked *

Scroll to Top