Text tool

PDF to Text

Choose a PDF, extract its embedded text page by page, and download a UTF-8 .txt file without uploading the document.

In-browser processingNo account requiredPrivacy details ↗
No file selected

One PDF up to 50 MB. The file stays in this browser and is never uploaded.

The result is capped at 1,000,000 characters and 64 MB estimated text memory.

Choose a PDF to begin.

A QUICK WALKTHROUGH

How to use this tool

  1. Choose one local PDF up to 50 MB and set an optional page range.
  2. Extract the embedded text with PDF.js, keeping page breaks and checking the output limits.
  3. Review the text, then download the UTF-8 .txt file. Scanned pages need OCR elsewhere.

Embedded text only

This tool reads text objects already present in the PDF with PDF.js. It does not OCR scanned pages, inspect images, or claim that a page is empty when its words are only pixels. A clear no-text warning helps you choose an OCR tool when needed.

Bounded local processing

The PDF stays on this device. The page limits the file size, selected page range, output characters, and estimated in-memory text so a very large document does not silently exhaust the browser. Text is separated with page markers and downloaded as UTF-8.

GOOD TO KNOW

Common questions

Does this perform OCR?

No. It extracts embedded PDF text only. Scanned or image-only pages need a separate OCR workflow.

Can I extract selected pages?

Yes. Enter a start and end page, or leave the range empty to process the whole PDF.

Is my PDF uploaded?

No. The selected local file is opened by PDF.js in this browser and is never sent to a server.

Why did extraction fail?

The PDF may be damaged, password-protected, unsupported, or too large for the stated browser limits.