Skip to content
Pdfiqo

Extract text from a PDF

Copy all the text out of a PDF as a .txt file.

Drop a PDF hereYour files stay on this device.Choose files
  • Processed on your device
  • Works offline once loaded
  • No sign-up, no watermark

How to pDF to Text

  1. 1

    Load the PDF

    Drop a PDF of up to 200 MB onto the page or browse to it. Encrypted files will ask for their password before anything is read.

  2. 2

    Decide on line breaks

    Leave Keep visual line breaks off for flowing text that is easy to paste elsewhere. Switch it on to rebuild lines based on where they sit on the page.

  3. 3

    Extract the text

    Press PDF to Text. The words are pulled out page by page, with a clear marker such as --- Page 2 --- between pages.

  4. 4

    Save the .txt file

    Click Download to keep a plain text file named after your original document. It opens in any editor on any operating system.

Extract text from any PDF

A PDF is great for reading but awkward when you want to reuse its words. Copying and pasting from a viewer often breaks lines mid-sentence, skips columns or runs out of patience on long files. Extracting the text into a plain .txt file gives you the raw content in one go, ready for search, analysis, quoting, translation or feeding into another tool.

This extractor walks through every page, collects the text layer and writes it to a single UTF-8 file. Each page begins with a simple header like --- Page 1 --- so you can always tell where something came from, which is especially helpful when citing reports or checking long contracts.

Flowing text versus visual lines

The one option, Keep visual line breaks, changes how the words are arranged.

  • Off produces text that reads like normal prose. It is the better choice for pasting into an email, a word processor or a translation service.
  • On rebuilds lines from their coordinates on the page. Use it when the arrangement matters, such as for invoices, forms, poetry or code listings where each line carries meaning.

If one mode gives awkward results for a particular document, simply run it again with the other setting. Processing is quick, so experimenting costs nothing.

Everything stays in your browser

Your file is not sent anywhere to be parsed. The PDF engine is loaded into the web page itself, reads your document from local memory and writes the .txt file on the spot. That is a meaningful difference for anyone working with legal drafts, HR records, research data or client correspondence, because the content is never exposed to a third-party server, log file or storage bucket.

Limitations to know about

Text extraction can only return text that genuinely exists in the file. Scanned pages, photographed receipts and image-only PDFs contain pictures of letters rather than characters, so they come out blank. The fix is to make them searchable first with OCR PDF and then extract.

A few other honest caveats apply. Multi-column layouts may interleave in unexpected ways, text that is drawn as vector outlines cannot be read, and tables lose their grid, leaving cells separated by spaces. Formatting such as bold, italics and font sizes is not part of a plain text file.

Where to go next

For an editable document that keeps paragraphs and headings, PDF to Word is the richer option. Need to turn a text file back into a clean, paginated document? Text to PDF handles that. And if you actually want pictures of the pages rather than their words, PDF to JPG renders them as images.

Frequently asked questions

How do I extract text from a PDF for free?

Add your file here and press PDF to Text. You receive a plain .txt file containing every page of text, with no account, watermark or fee involved.

Why is the extracted text empty or full of gaps?

Your PDF is probably a scan, meaning each page is a photo rather than real text. Run it through OCR PDF first to recognize the words, then extract the text from the searchable result.

What does Keep visual line breaks do?

With it on, lines are reconstructed from their position on the page, so the output mirrors the visual layout more closely. With it off, you get cleaner running text that is easier to reflow in another document.

Does PDF to Text keep formatting like bold, fonts or tables?

No. A .txt file holds characters only, so styling, images and table borders are dropped. If you need an editable document with headings and paragraphs, try PDF to Word.

Is my document uploaded when I extract its text?

It is not. Text extraction runs entirely in the browser on your computer or phone, and the resulting file is generated locally.

Can I extract text from a password-protected PDF?

Yes, as long as you know the password. You will be prompted for it, and the file is decrypted in memory on your device for the extraction.

Which languages are supported?

Any language whose text is actually embedded in the PDF comes out as Unicode, including accented Latin letters, Cyrillic and Asian scripts. The output is saved as UTF-8, which modern editors read without trouble.