The quickest way to convert PDF to text is to open the PDFNova PDF to Text tool, add your file, and either copy the extracted words or save them as a .txt file. It works on digital PDFs, the kind made by exporting from a word processor or a website. Scanned PDFs are pictures of pages, so they need OCR first.
Plain text is the most portable form a document can take. It opens on every device, pastes cleanly into any app, and carries no hidden formatting.
Reasons to turn a PDF into plain text
Copying directly from a PDF viewer is fine for a sentence. For anything longer it becomes slow and messy, especially on a phone. Extracting the whole text at once is better when you want to:
- quote or summarise a long report, paper or contract
- paste content into an email, a chat, a form or a note-taking app
- translate a document or run it through a spell-checker
- count words or search for terms in a simple text editor
- reuse the wording in your own template without bringing old fonts and spacing along
- keep a small, lightweight copy of the content for reference
How to convert PDF to text with PDFNova
- Open PDF to Text in your browser on a phone, tablet or computer.
- Add the PDF. If the file is password-protected, enter its password when asked.
- Wait while the tool reads the text from the pages.
- Review the extracted text on screen.
- Tap copy to place it on your clipboard, or save it as a .txt file to your device.
PDFNova processes files in your browser, so your documents are not uploaded to a server. It is free, needs no sign-up and the default maximum file size is 100 MB.
Digital PDF or scanned PDF: pick the right tool
This is the single most important thing to check. Open the PDF and try to highlight a few words. If they highlight, the PDF is digital. If the page behaves like one big image, it is a scan.
| Your file | Use this | What you get |
|---|---|---|
| Digital PDF with selectable text | PDF to Text | The exact characters stored in the file, to copy or save as .txt |
| Scanned PDF or photo of a page | OCR PDF | Recognised text to copy, a .txt file, or a searchable PDF |
| Digital PDF you want to keep editing with headings | PDF to Word | An editable .docx with headings and paragraphs detected |
Extraction from a digital PDF copies characters that already exist, so it does not misread letters. OCR has to recognise shapes, so its accuracy depends on how clean the scan is. If you have a choice between a digital copy and a scan of the same document, always use the digital one.
What plain text keeps and what it drops
A .txt file holds characters and line breaks. Nothing else. Knowing this in advance avoids surprises.
- Kept: the words, numbers, punctuation and basic line breaks.
- Dropped: fonts, sizes, bold and italics, colours, images, charts, page layout.
- Changed: tables become rows of text, and columns may run together.
If you need structure such as headings and paragraphs in an editable document, PDF to Word is the better fit. If you need the pictures, extract those separately.
Troubleshooting messy or missing text
The result is empty
The PDF is almost certainly a scan with no text inside. Use OCR instead. Some PDFs are also mixed, with typed pages and scanned pages together, so only part of the text appears.
Every line ends too early
PDFs store text line by line as it appears on the page. When extracted, each printed line can become its own line. Paste the text into an editor and join the lines of each paragraph, or use find and replace to remove single line breaks while keeping the blank lines between paragraphs.
Words are joined together or split apart
Some PDFs position letters individually without real spaces. The extracted text reflects what is stored. A quick spell-check finds most of these.
Text from two columns is mixed
The reading order in a PDF does not always match what the eye sees. For multi-column pages, expect to rearrange sections after extracting.
Headers, footers and page numbers appear in the middle of the text
They are part of each page, so they are extracted with it. Delete the repeated lines. If the same header appears on every page, find and replace clears them quickly.
Urdu or Arabic text comes out scrambled
Right-to-left scripts are often stored in a PDF in a way that suits display, not copying. Letters can appear reversed, disconnected or in the wrong order. We explain the causes and the workarounds in why Urdu text breaks when copied from a PDF and how to fix it. In many cases, running OCR with the Urdu or Arabic language selected gives more usable text than direct extraction.
After extracting: clean-up checklist
- Compare the first and last lines with the PDF to confirm nothing was cut off.
- Check figures, dates and names against the original.
- Remove repeated headers, footers and page numbers.
- Rejoin broken paragraphs.
- Save the file with a clear name so you can find it again.
Going the other way: text back into a PDF
Once you have edited the text, you may want a tidy document again. The Text and HTML to PDF tool turns pasted text or a .txt file into a paginated PDF. Our walkthrough on how to convert a TXT file to PDF covers the steps. If you want headings, bold text or lists in the result, simple HTML gives you that control, as described in how to convert HTML to PDF.
Frequently asked questions
How do I copy all the text from a PDF at once?
Add the file to PDF to Text and use the copy option. The full text of the document goes to your clipboard, ready to paste anywhere.
Why can I not copy text from my PDF?
There are two common reasons. The PDF may be a scan, which has no text to copy and needs OCR. Or the author may have set a restriction on copying, in which case you should ask them for a copy you are allowed to use.
Does converting a PDF to text keep the tables?
The contents are kept but the grid is not. Each row becomes a line of text. For data you need to reuse, check the numbers carefully and rebuild the table in a spreadsheet or document.
Can I extract text from a PDF on a phone?
Yes. The tool runs in a modern mobile browser, and copying to the clipboard makes it easy to paste the text into a message or notes app.
For digital PDFs, extraction is fast and exact. For scans, use OCR and proofread. Open the free PDF to Text tool to pull the words out of your document now.