VSThiran

How to convert a PDF to plain text

Quick answer

Upload your PDF, click Extract text, and copy the plain text directly or download it as a .txt file. Real, selectable text converts directly; a scanned PDF has none to extract and needs OCR first.

  1. 1Upload your PDF.
  2. 2Click Extract text.
  3. 3Copy the text, or download it as .txt.

Free, no sign-up, and your file is read on your own device rather than uploaded.

Sometimes you do not want a Word document or a spreadsheet - you just want the actual words in a PDF, as plain text, to paste into an email, search through, or feed into something else that only wants raw content.

Whether that works depends on what kind of PDF you have. One built from a Word document, exported from a webpage, or saved from an app already has a real text layer - it is "searchable", meaning you can Ctrl+F through it, and this tool reads that text directly in your browser. A scanned PDF - a photograph or scanned image of a page - has no text layer at all, just a picture of words, so there is nothing to extract until OCR creates one.

"Searchable" is not the same as "editable", which is a common mix-up. A searchable PDF's text is real, but it is locked inside a fixed page layout. Extracting it as plain text, here, throws that layout away entirely and gives you back just the words; keeping a layout you can actually type into and reflow is a different job, done by PDF to Word.

Step by step

  1. Upload your PDF

    Read directly in your browser - you are told the page count and roughly how many words it contains.

    PDF to Text
  2. Extract text

    The document is read and the text pulled out in reading order.

    PDF to Text
  3. Copy or download

    Copy the text directly to your clipboard, or download it as a .txt file.

    PDF to Text

Tips

  • If the PDF is a scan - a photograph or scanned image of a page - there is no real text in it to extract, and this tool will say so rather than returning nothing with no explanation. Run the OCR workspace on it first.
  • For a document with headings, lists or tables you want to keep as structure rather than flatten into plain text, the same tool also offers a Markdown view.
  • Plain text strips everything but the words - if you need the layout roughly preserved for editing, PDF to Word is the better fit.

Common problems

The extracted text came back empty.

The PDF is almost certainly a scan - a picture of a page rather than real text. Check with the PDF Inspector, then run the OCR workspace on it to get real text out.

Words from different columns are jumbled together.

A multi-column layout is read in a single flow by position, which can interleave columns on a complex page. There is no column-aware reading order in plain text extraction.

Examples

A contract exported from Word

You export a signed contract to PDF and need the text for an email. It is already searchable, so Extract text pulls the words out directly in a few seconds, ready to paste.

A scanned invoice from a supplier

A supplier emails a PDF that is actually a photo of a printed invoice. Extract text comes back empty because there is no text layer to read - run the OCR workspace on it first, then extract the real text it creates.

A two-column research paper

A PDF laid out in two columns is read top-to-bottom by position, not column by column, so the extracted text can interleave the left and right columns on a busy page. The Markdown view handles headings and paragraph breaks more predictably if that matters for your document.

Frequently asked questions

How do I get just the text out of a PDF?
Upload it to the PDF to Text tool and click Extract text - you get the words in reading order, with no formatting, ready to copy or download as .txt.
Why is the result empty?
The PDF is almost certainly a scan with no real text to extract - check with the PDF Inspector, then run OCR on it to recover the text.
What is the difference between a searchable PDF and an editable one?
A searchable PDF has a real text layer, so you can select, search and extract its words - that is what this tool uses. An editable PDF additionally needs its layout turned into something you can type into and reflow, which is a different job handled by PDF to Word.
Can I get Markdown instead of plain text?
Yes - the same tool offers a Markdown view that keeps headings, lists and detected tables as real structure, alongside the plain text.
Is my PDF uploaded to a server?
No. The file is read and converted entirely in your browser.

Doing several things to this PDF?

"Editing a PDF" is usually one of five different jobs. Open the file once and you can do all of them in sequence - reorder the pages, mark it, number it, sign it and shrink it - then download once at the end.

How to edit a PDF without Acrobat

Related PDF tools

Related articles

Ready to do it?

Free, no sign-up, and nothing is uploaded to a server.