PDF to TXT
Pull the text out of a PDF page by page and rebuild lines and paragraphs the way they were laid out. Export as TXT, Markdown or Word, pick a page range, merge wrapped lines and bundle pages into a ZIP - everything runs locally in your browser.
Choose a PDF file
Click to choose or drag a PDF hereUp to 50MB and 300 pages per file; nothing is ever uploaded
Extraction settings
Output format
Layout recovery
By coordinates looks at where every glyph sits, so columns and alignment survive. Content order follows the order the PDF was written in, which helps when coordinates are scrambled.
Line handling
Lines in a PDF are often wrapped mid-sentence. Merge CJK lines only joins Chinese neighbours; merge into paragraphs flattens each page into one paragraph and adds spaces between English words.
Page marks
Leave empty for every page. Comma lists, ranges and open-ended tails are supported.
Extracted text
Upload a PDF and the extracted text shows up here
How to use
- Upload a PDF (click or drag). A sample PDF is loaded on first paint, so you see results without clicking anything.
- Fill in a page range for just the pages you need: 1-3,5,8- means pages 1 to 3, page 5, and page 8 to the end.
- Chinese PDFs often wrap one sentence across several lines - switch line handling to "Merge CJK lines" to stitch them back together.
- "By coordinates" suits columns and tables; switch to "By content order" when a PDF has scrambled coordinates.
- The Word export produces a .doc that Word opens directly, paragraphs and line breaks intact. Scanned PDFs have no text layer and need OCR first.