PDF to Markdown
Extract a PDF's text as clean Markdown — headings detected from font size, bullets converted, page breaks preserved. Perfect for LLM/RAG ingestion, notes, and docs. Runs entirely in your browser.
Drop a PDF here
or click to browse — text-layer PDFs (scans need OCR)
Why markdown, and honest limits
- • Headings are inferred from relative font size (≥70% larger = h1, ≥40% = h2) — check important docs
- • Bullet glyphs (•, ▪, -, ·) become markdown list items
- • Tables come out as text — complex layouts flatten; this extracts the text layer, not the layout
- • Scanned PDFs have no text layer — use OCR first, then convert
- • Related: PDF to Text · HTML to Markdown