Convert PDF to Markdown
PDF to Markdown turns the existing text of a PDF into a .md file, with headings, paragraphs and lists where it can detect them. It does not OCR, so scanned pages come out empty.
Drag and drop files here, or click to choose
Accepts PDF · up to 500.0 MB per file · up to 20 files at once
You can also paste a file with Ctrl/Cmd+V, or from the clipboard menu.
PDF to Markdown runs pdf-inspector on the file's existing text and returns a `.md` download (`text/markdown`). Headings, paragraphs, and lists are the goal. Pixel-perfect layout is not.
It does not OCR. Image-only pages often come out empty. For scans, use OCR to Markdown; that tool already returns `.md` and is the one that looks at pixels.
Reading order can disagree with a multi-column visual layout. Running headers may repeat every page; delete them in an editor. Equations that were images in the PDF do not become LaTeX. Code listings often lose indentation.
This is the converter for docs-as-code, Git diffs, and pipelines that want text. Keep the PDF when print layout is the record. Copyright still applies to the extracted words; only convert files you may republish as text.
Password-protected files are unlocked for you on this page once you enter the password; API callers run Unlock first. Open the Markdown and fix heading levels before you commit it. Markdown to PDF on this site is the reverse direction; a round trip will not restore the original layout.
Uploads are temporary. We do not keep a Markdown library of your documents.
Features
- pdf-inspector extract of native PDF text to Markdown
- Does not OCR; scans need the OCR tool
- Download is `text/markdown` with a `.md` name
- Anonymous API at /api/v1/convert/pdf/markdown
When to use this tool
- Move a PDF README into a Git repository
- Feed article text into a prompt pipeline
- Draft web copy from a PDF brief
How do I convert a PDF to Markdown?
- Prefer a text-based PDF. For scans, use OCR to Markdown instead.
- Upload the PDF and click Process.
- Download the `.md` file.
- Fix heading levels, repeated headers, and tables in your editor.
Limits and edge cases
- No OCR; image-only pages extract little or nothing
- Layout fidelity is limited
- Complex tables may need hand fixes
- Upload limit on this website: 500 MB per file, sent in chunks above about 95 MB. A single direct API request body is capped at 100 MB.
Examples
- A text whitepaper becomes a .md file with paragraphs and headings
- A scan without OCR yields little or no useful Markdown
Privacy for this tool
Files are uploaded over HTTPS, processed in memory or a short-lived temporary directory on our servers, and deleted when your result is ready. We do not keep copies for later browsing, training, or advertising profiles. See the Privacy Policy for retention details and AdSense cookie disclosures.
Frequently asked questions
- Why is my Markdown empty?
- The PDF likely has no text layer. Use OCR to Markdown for scans. This tool only reads text already in the file.
- Is this the same as OCR?
- No. OCR looks at image pages. PDF to Markdown uses pdf-inspector on native text.
- Does this keep fonts?
- Markdown is text plus light structure. Fonts come from whatever renders the Markdown later.
- Can I convert Markdown back to PDF?
- Yes via Markdown to PDF. Layout will not match the original PDF.
Last updated:
Call this from code
Every tool on this site is a plain REST endpoint - no account or API key needed for anonymous use. Built for AI agents and developers as much as for browsers.
curl -X POST "https://pdf123.xyz/api/v1/convert/pdf/markdown" \
-F "[email protected]" \
-o output.pdfAlso available as an MCP tool for agent clients that speak Model Context Protocol (JSON-RPC 2.0 over POST /mcp). Full API reference
Use it from an AI agent
SkillClaude Code, Codex, Cursor and other AI agents can run PDF To Markdown for you with this skill: /skills/pdf123-pdf-to-markdown.md
Files are used only for this processing job and deleted automatically afterward.