PDF to ODS (Text Transcript)
Create a real ODS workbook for sorting and searching page transcripts. This does not detect or reconstruct PDF tables.
LOCAL TEXT TRANSCRIPT
Extract PDF text into ODS
Create one worksheet row per page. This is a text transcript, not table recognition.
Embedded text only. Scanned pages receive a no-text marker; OCR is not provided. The tool does not identify tables, cells, columns, formulas, images, or reading order.
PDF parsing and ODS generation happen in this browser. Your file is never uploaded.
A QUICK WALKTHROUGH
How to use this tool
- Choose a PDF up to 50 MB and 100 pages.
- Extract existing text objects locally, one page at a time.
- Download an ODS workbook with page and text columns. Scanned pages are marked; OCR is not performed.
A real OpenDocument spreadsheet
The downloaded ODS is a valid OpenDocument Spreadsheet package with a Page column and a Text transcript column. It is useful for searching or sorting extracted page text, but it does not recognize tables, cells, formulas, or document layout.
Embedded text only
PDF.js reads existing text objects in their available order. Image-only pages receive a no-text marker; OCR is not performed. Multi-column order, tables, footnotes, and semantic structure are not reconstructed.
Private and bounded
PDF parsing and spreadsheet generation happen in your browser; the PDF is never uploaded. One PDF may be up to 50 MB and 100 pages, extracted text up to 25 million characters, and output up to 100 MB.
GOOD TO KNOW
Common questions
Does this recognize PDF tables?
No. It creates one row per page with extracted text. It does not detect table rows, columns, formulas, or cells.
Does it read scanned pages?
No. It reads embedded text only and marks pages without extractable text. OCR is not included.
Is my PDF uploaded?
No. PDF parsing and ODS generation run locally in your browser.