My Tool Studio
PDF

PDF to Markdown Converter

This PDF to Markdown converter turns a PDF into clean Markdown you can paste into a README, a notes app, a static site or an AI chat. Headings are detected from font sizes and bold lines, and lists, tables made of text in columns, bold, italic, inline code and links become proper Markdown syntax. Running headers, footers and page numbers can be removed, and words split across lines are joined again. Scanned pages are read with OCR, and pictures can be saved as PNG files next to the .md in a ZIP. Preview the rendered result, copy it, or download it. Several PDFs can be converted at once, and nothing is uploaded. Heading levels are an estimate, so check them in long documents.

Always freeNo sign upRuns in your browser

Leave empty for all, or type ranges like 1-3, 5.

Scanned pages have no text layer. OCR reads them from the page image.

The first run downloads this language's data.

For pages that mix two languages, such as Hindi and English.

Good to know

  • Headings are guessed from font size and bold lines, so check the levels in long documents.
  • Tables are found from text lined up in columns. Merged cells and tables drawn as pictures are not detected.
  • Charts, drawings and text inside pictures are not converted. OCR only reads scanned pages.
  • Footnotes stay where they appear on the page.

How to use

01

Add one or more PDFs

Drop PDFs of up to 200 MB each, reorder them if needed, and type a page range such as 3-12 to skip covers or appendices.

02

Choose what to keep

Pick a page marker (nothing, a rule, or an HTML comment), the OCR mode, and tick headings, tables, formatting, links, pictures, header removal and hyphen repair.

03

Convert, check and copy

Press Convert to Markdown. Switch between the Markdown and Preview views, then press Copy or download the .md file, or a ZIP when pictures are included.

Why PDF to Markdown Converter

How PDF content maps to Markdown

Each element found on the page becomes a standard Markdown construct.

In the PDFIn the Markdown
Larger text sizes# to ### headings
Short standalone bold linesA lower heading level
Bullets and numbered items- items and 1. items, nested by indent
Columns of aligned text| table | rows |
Bold, italic**bold**, *italic*
Monospaced text`inline code` or fenced code blocks
Links[text](url)
Pictures (optional)![Image from page N](images/...png)

Cleaning up for AI and search

Running headers, footers and lone page numbers repeat on every page and break paragraphs apart. With the removal option on, lines that repeat in the top or bottom band of most pages are dropped, and paragraphs that continue onto the next page are joined when no page marker is used.

Common questions

How do I convert a PDF to Markdown?
Drop the PDF and press Convert to Markdown. The Markdown appears in a box with Copy and Download buttons, and the Preview view shows how it renders.
Why convert a PDF to Markdown before using it with ChatGPT or Claude?
Markdown keeps the structure (headings, lists, tables) in plain text, which AI models read well, and removing repeated headers and page numbers cuts noise. It also makes the text easy to split into sections.
How does PDF to Markdown decide what is a heading?
It finds the most common body text size, then treats larger sizes as heading levels, biggest first. Short bold lines at body size that stand on their own become the next level down. Untick Detect headings to get plain paragraphs.
Does PDF to Markdown convert tables?
Yes, when text is lined up in columns across two or more rows. Those rows become a GitHub-style Markdown table with the first row as the header. Tables with merged cells or tables drawn as pictures are not detected.
Can I keep the images when converting a PDF to Markdown?
Tick Save pictures as PNG files. Each picture is cut from the page, saved in an images folder, and linked from the Markdown. The download becomes a ZIP with the .md file and the images folder.
Can I mark where each PDF page starts in the Markdown?
Yes. Set Between pages to Horizontal rule for a --- line, or HTML comment for an invisible <!-- Page N --> marker that keeps page numbers for reference.
Does PDF to Markdown work on scanned documents?
Yes. With Text recognition on Auto, pages without a text layer are read with OCR in the language you choose. OCR text has no bold or italic information, so only headings by size, lists and paragraphs are detected there.
Which Markdown flavour does this converter produce?
It writes CommonMark with GitHub-style tables, so the output works on GitHub, GitLab, Obsidian, Notion imports and most static site generators.
Is my PDF uploaded when converting to Markdown?
No. The PDF is read with pdf.js and the Markdown is written in your browser. Only the OCR language data is downloaded, the first time you use OCR.

More PDF tools

View all