Pull every word out of your PDF into a clean plain-text file — perfect for search indexing, analysis, quoting, and repurposing content.
Only if it contains a text layer. Pure image scans need OCR first — PolyForge extracts existing embedded text.
A plain UTF-8 .txt file (or a ZIP of per-page .txt files in per-page mode).
Text is extracted in the document’s logical page and flow order, which matches visual order for the vast majority of PDFs.