PDF to Text
Choose a PDF, extract its embedded text page by page, and download a UTF-8 .txt file without uploading the document.
A QUICK WALKTHROUGH
How to use this tool
- Choose one local PDF up to 50 MB and set an optional page range.
- Extract the embedded text with PDF.js, keeping page breaks and checking the output limits.
- Review the text, then download the UTF-8 .txt file. Scanned pages need OCR elsewhere.
Embedded text only
This tool reads text objects already present in the PDF with PDF.js. It does not OCR scanned pages, inspect images, or claim that a page is empty when its words are only pixels. A clear no-text warning helps you choose an OCR tool when needed.
Bounded local processing
The PDF stays on this device. The page limits the file size, selected page range, output characters, and estimated in-memory text so a very large document does not silently exhaust the browser. Text is separated with page markers and downloaded as UTF-8.
GOOD TO KNOW
Common questions
Does this perform OCR?
No. It extracts embedded PDF text only. Scanned or image-only pages need a separate OCR workflow.
Can I extract selected pages?
Yes. Enter a start and end page, or leave the range empty to process the whole PDF.
Is my PDF uploaded?
No. The selected local file is opened by PDF.js in this browser and is never sent to a server.
Why did extraction fail?
The PDF may be damaged, password-protected, unsupported, or too large for the stated browser limits.