convertCASEpro

MODE

PDF to Markdown Converter

Drop a PDF and get structured Markdown back.

How the PDF to Markdown Converter works

The tool reads the text layer of each page together with its position and font size. The most common font size is treated as body text, larger lines become headings (one to three levels, depending on how much larger), lines starting with a bullet or number become list items, and lines that are close together are merged into paragraphs. Words split by a hyphen at a line end are joined again.

This approach works for text based documents such as reports, articles and manuals. It cannot read scanned pages, which contain only pictures. For those, run the page images through the Image to Text Converter. Complex layouts such as multi column pages and tables will need some manual editing.

When it is useful

  • •Moving a report or paper into a note taking app or a wiki.
  • •Extracting readable text from a PDF to feed into other tools.
  • •Preparing documentation PDFs for a static site.
  • •Getting an editable draft from a PDF whose source file is lost.

Frequently Asked Questions

Is my PDF uploaded?

No. The file is read in your browser with pdf.js and never sent to a server.

Why is the result empty?

The PDF probably contains scanned images instead of text. Use the Image to Text Converter on the pages instead.

Are tables converted?

Not as tables. Their cell text appears as lines. Rebuild them with the Markdown Table Generator if needed.

Related tools

[ Browse all tools ]