PDF to Markdown

Extract a PDF's text as clean Markdown — headings detected from font size, bullets converted, page breaks preserved. Perfect for LLM/RAG ingestion, notes, and docs. Runs entirely in your browser.

Drop a PDF here

or click to browse — text-layer PDFs (scans need OCR)

Why markdown, and honest limits

  • • Headings are inferred from relative font size (≥70% larger = h1, ≥40% = h2) — check important docs
  • • Bullet glyphs (•, ▪, -, ·) become markdown list items
  • • Tables come out as text — complex layouts flatten; this extracts the text layer, not the layout
  • • Scanned PDFs have no text layer — use OCR first, then convert
  • • Related: PDF to Text · HTML to Markdown