VSThiran

How to convert a PDF to Markdown

Quick answer

Upload your PDF, click Extract text, and switch to the Markdown view. Headings come from font size, list items are recognised from bullets or numbers, and tables with clean columns become real Markdown tables - then download as .md.

  1. 1Upload your PDF.
  2. 2Click Extract text.
  3. 3Switch to the Markdown view.
  4. 4Download as .md.

Free, no sign-up, and your file is read on your own device rather than uploaded.

Markdown is what most documentation - a README, a wiki page, a static site, a lot of note-taking tools - actually reads as formatting rather than literal text, so a PDF that needs to go into one of those needs to become Markdown, not just plain text.

This reconstructs headings, bullet and numbered lists, and tables directly from the PDF - the same font-size and table-detection heuristics PDF to Word already uses, just written out as Markdown syntax instead of a Word document.

Step by step

  1. Upload your PDF

    Read directly in your browser.

    PDF to Text
  2. Extract text

    The document is read once, structure and all.

    PDF to Text
  3. Switch to the Markdown view

    Headings, lists and detected tables appear as real Markdown syntax.

    PDF to Text
  4. Copy or download as .md

    Copy directly, or download the Markdown file.

    PDF to Text

Tips

  • Headings are inferred from font size relative to the rest of the page - a document with a clear, consistent heading hierarchy converts more accurately than one where headings and body text are similar sizes.
  • A table needs genuinely aligned columns to become a Markdown pipe table; text that only looks aligned by eye will come through as plain lines instead.
  • Markdown has no way to represent a multi-column page layout, exact positioning, or images pulled from the PDF - only the text structure survives.

Common problems

A heading was not detected as a heading.

Heading detection relies on font size standing out from the body text - if a heading is only slightly larger, or styled with bold rather than size, it may come through as a plain paragraph instead.

A table came through as plain lines instead of a Markdown table.

The original table probably lacks true column alignment - text nudged into place with spacing rather than genuine columns does not convert into a reliable table.

The PDF is a scan and produced no Markdown at all.

A scan has no real text to structure into Markdown. Run the OCR workspace on it first, then bring the extracted text back here.

Frequently asked questions

Does it keep headings and lists as real Markdown?
Yes - heading levels come from font size relative to the page, and lines starting with a bullet or number are converted into Markdown list syntax.
Are tables converted to Markdown tables?
Where the table has genuine column alignment, yes - a Markdown pipe table. Where it does not, the rows come through as plain text instead.
Does it work on a scanned PDF?
Not directly - a scan has no text to structure. Run the OCR workspace on it first, then convert the extracted text here.
Can I get plain text instead of Markdown from the same file?
Yes - the same tool offers a plain text view alongside Markdown, from the same extraction.
Is my PDF uploaded anywhere?
No. Conversion happens entirely in your browser.

Related PDF tools

Related articles

Ready to do it?

Free, no sign-up, and nothing is uploaded to a server.