Drop a PDF and get structured Markdown back.
The tool reads the text layer of each page together with its position and font size. The most common font size is treated as body text, larger lines become headings (one to three levels, depending on how much larger), lines starting with a bullet or number become list items, and lines that are close together are merged into paragraphs. Words split by a hyphen at a line end are joined again.
This approach works for text based documents such as reports, articles and manuals. It cannot read scanned pages, which contain only pictures. For those, run the page images through the Image to Text Converter. Complex layouts such as multi column pages and tables will need some manual editing.
No. The file is read in your browser with pdf.js and never sent to a server.
The PDF probably contains scanned images instead of text. Use the Image to Text Converter on the pages instead.
Not as tables. Their cell text appears as lines. Rebuild them with the Markdown Table Generator if needed.