PDF to Text
在浏览器中使用「PDF to Text」完成文字与文档相关任务;输入内容和结果默认在设备上处理。
使用步骤
如何使用
- Choose one local PDF up to 50 MB and set an optional page range.
- Extract the embedded text with PDF.js, keeping page breaks and checking the output limits.
- Review the text, then download the UTF-8 .txt file. Scanned pages need OCR elsewhere.
Embedded text only
This tool reads text objects already present in the PDF with PDF.js. It does not OCR scanned pages, inspect images, or claim that a page is empty when its words are only pixels. A clear no-text warning helps you choose an OCR tool when needed.
Bounded local processing
The PDF stays on this device. The page limits the file size, selected page range, output characters, and estimated in-memory text so a very large document does not silently exhaust the browser. Text is separated with page markers and downloaded as UTF-8.
常见问题
你可能还想知道
Does this perform OCR?
No. It extracts embedded PDF text only. Scanned or image-only pages need a separate OCR workflow.
Can I extract selected pages?
Yes. Enter a start and end page, or leave the range empty to process the whole PDF.
Is my PDF uploaded?
No. The selected local file is opened by PDF.js in this browser and is never sent to a server.
Why did extraction fail?
The PDF may be damaged, password-protected, unsupported, or too large for the stated browser limits.