{"repo":"PSPDFKit/pdf-to-markdown","free":true,"listed":false,"github":"https://github.com/PSPDFKit/pdf-to-markdown","clone":"git clone https://github.com/PSPDFKit/pdf-to-markdown.git","description":"Standalone CLI wrapper and docs for Nutrient's PDF-to-Markdown extractor","language":"Shell","stars":317,"topics":["cli","markdown","nutrient","pdf","pspdfkit","parser","pdf-converter","pdf-parser","pdf2md","skill-md"],"license":null,"category":"media-processing","readme_excerpt":"Nutrient PDF to Markdown Stop wasting your context window on PDF extraction. Fast, accurate Markdown from PDFs — locally, with no cleanup required. Built for Claude, Codex, RAG pipelines, and document-heavy automation where noisy extraction burns tokens and makes downstream results less reliable. - How fast is it? — 0.004s per page. 134x faster than docling, 53x faster than pymupdf4llm. (benchmarks) - How accurate is it? — 0.93 reading order (best in class), 0.89 overall extraction accuracy, 0.82 heading detection. (benchmarks) - NEW: --vision tier — a licensed machine-vision ICR pipeline that tops every accuracy metric, including tables (0.94 TEDS), handles scanned and handwritten documents — and still runs faster than docling (0.35s per page). (benchmarks) - Three tools, one binary — pdf-to-markdown for structured Markdown, pdf-to-text for layout-preserving plain text, and query for ranked search over an extracted file. Pick by what the downstream consumer needs. (the Nutrient document CLI) - NEW: Image export — --enable-image-export extracts images alongside Markdown for vision-capable LLMs. (usage) - Where do my PDFs go? — Nowhere. The CLI runs locally. Your documents are not uploaded to Nutrient. (trust & licensing) - What does it cost? — Free for up to 1,000 documents per calendar month. No license key, no signup, no API token. (license) The Nutrient document CLI pdf-to-markdown is one verb of a single signed binary that turns digital-born PDFs into agent-ready output —","default_branch":null,"files":null,"tree":[],"storefront":"/r/PSPDFKit","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/PSPDFKit/pdf-to-markdown/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}