Skip to main content
ToolsBay

Runs in your browser — nothing is uploaded · verify

PDF to Markdown — free, and nothing is uploaded

Convert a PDF into Markdown with headings and lists detected.

  • Internal documentation stays internal: the PDF is parsed in the page, never uploaded.
  • Ready to paste into a wiki, a README, a static site or a note-taking app.
  • Font sizes and indentation drive the structure, so the output has real headings.

Drop the PDF you want as Markdown

or click to browse · .pdf · up to 100 MB

  • 70 tools

    Free, with no account and no watermark

  • 0 bytes

    Of your files or text uploaded — this tool runs in the tab

  • No queue

    The work happens in this tab, not in a line behind other users

Three steps, nothing to install

PDF to Markdown in three steps

  1. PDFDrag it in, or click to browse

    Add the PDF

    Drop it in. Parsing happens in the page, so internal documentation stays internal.

  2. PDFMARKDOWNConverted on your own device

    Detect the structure

    Relative text size marks the headings and leading characters mark the list items, so the Markdown has real structure rather than a flat wall of paragraphs.

  3. Download MARKDOWN

    Take the Markdown

    Copy it straight out or download the .md file, ready to commit.

Why use PDF to Markdown

  • Headings, not just big text

    Type size is compared across the whole document to work out the heading levels, which is what makes the result usable as a document rather than as a transcript.

  • Headings and lists detected

    Lines set larger than the body text become # headings, and leading bullets and numbering become list items. Either can be switched off.

  • The file never leaves the tab

    There is no upload because there is nowhere to upload to. The site is static files, and the work runs as code on your own machine.

Frequently asked questions

Markdown keeps the structure that plain text throws away. Headings stay headings and lists stay lists, so the document is still navigable — and a language model reading it can tell a section title from a sentence. For a straight dump with no structure at all, use PDF to Text instead.

PDF to Markdown, without uploading the document

Markdown has quietly become the format that machines and humans agree on. It is what static site generators build from, what wikis and note apps store, and — increasingly the reason people want this conversion — what language models read most reliably. A PDF pasted into a chat box arrives as a jumble; the same document as Markdown arrives with its structure intact.

The catch is that the documents worth feeding to an AI assistant are usually internal: a contract, a specification, a research paper under licence, a board pack. Uploading those to a conversion service to prepare them for analysis defeats the purpose. This tool runs the extraction in your browser, so the file never leaves your device.

What the converter reconstructs

A PDF does not store paragraphs, headings or lists. It stores glyphs at coordinates — which is why copying text out of one so often produces a mess. Rebuilding structure means inferring it from geometry.

Lines are grouped from runs sharing a baseline, with the tolerance scaled to glyph height so a large heading is not split in two. Paragraphs are separated where the vertical gap grows beyond normal line spacing. Headings are lines set larger than the document's body size, with the size ladder mapped onto heading levels. Lists are recognised from leading bullet glyphs and numbering, and re-emitted as Markdown list syntax. Characters that would otherwise be read as Markdown markup are escaped, so a literal asterisk in the source stays an asterisk.

Where it stops

Tables come through as plain lines, because most PDFs give no indication that a group of text runs is a grid. Multi-column layouts — academic papers especially — are read in the order the file stores them, which is usually column by column but is not guaranteed. Images and equations are not extracted. And a scanned document has no text layer at all, so there is nothing to convert; the tool reports that rather than returning an empty file.

These are limits of how most PDFs are written as much as of this implementation; a PDF tagged for accessibility can carry more structure, which this tool does not read. Knowing where the output needs checking is more useful than a promise it will be perfect.

Related tools

For an unstructured dump, PDF to Text is simpler. For an editable document, use PDF to Word. For a styled web page, PDF to HTML. To count what you extracted, try the Word Counter.

  • PDF to Text

    Extract the plain text content of a PDF.

  • PDF to HTML

    Convert the text layout of a PDF into an HTML document.

  • PDF to Word

    Extract the text of a PDF into an editable Word (DOCX) document.

  • Merge PDF

    Combine several PDFs into one, in the order you choose.

  • Split PDF

    Split a PDF by page range, every N pages, or one file per page.

  • Organize PDF

    Rotate, reorder and delete PDF pages using visual thumbnails.

More tools

70 tools, none of which want your file

Everything this tool does happens in the page you are looking at. No account, no upload, no watermark.