Convert PDF to Markdown locally: extract its text layer or run OCR on selected pages without text. Review and edit before copying or downloading.
What is PDF to Markdown?
PDF to Markdown is a free online tool that extracts the text of a PDF to editable Markdown in your browser. It reads the text layer of the file page by page, shows the result for you to correct and tells you which pages have no text.
What it supports
- PDF text-layer extraction, plus optional local OCR on selected pages without a text layer in English, Portuguese or Spanish.
- Text extracted page by page, as editable Markdown, in approximate reading order.
- Copy the result or download it as an MD file.
Good to know
- Reading order is approximate, and headings, columns and tables are not reliably rebuilt, so expect to edit the result.
- No text layer can mean a scan or a blank page; it does not prove a page is scanned. OCR is optional and approximate. It replaces the generated Markdown, including manual edits, so run it before correcting the result.
- PDF scripts and actions are not run.
How do I use PDF to Markdown?
- Use Open file or drop a PDF into PDF file. Text extraction starts automatically.
- For pages marked No text layer, select up to 10, choose Recognition language and press Run OCR on these pages. Check progress or use Cancel.
- Edit the Markdown, then copy or download it.
Example
A PDF page with selectable text.
[PDF page with a text layer]
Quarterly report
Revenue grew 8% in March.The text of the page comes out as Markdown to edit. Headings and layout are not rebuilt, and a page without text would be flagged.
Limits and privacy
Limits
- File: up to 20,000,000 bytes (about 20 MB). Larger or unreadable files are rejected, and the text taken from a file is also limited to 2,000,000 bytes.
- PDF: up to 100 pages. A longer PDF is rejected.
- OCR: select 1 to 10 pages without a text layer per run. Each rendered page is capped at 16,000,000 pixels and 16,384 pixels per side. Recognition stops after 120 seconds.
Is my content sent to a server?
No. Your content and files are processed in your browser and never uploaded. OCR and PDF reading download processing files, including OCR language data and the PDF worker, from this same site only when those tools run; no content goes to third parties. Markdown previews show external images as alt text instead of loading them.
Copying puts a result on your clipboard and downloading saves it as a file, but only when you press the button. The file is created in your browser, so nothing is uploaded.
For the full details, see the Privacy Policy
Related tools and pages
Convert to or from other formats, or work on the same content with these tools.