PDF to HTML
Keep the visual appearance of each PDF page in a portable HTML file, with any existing selectable text included as a copyable transcript.
A QUICK WALKTHROUGH
How to use this tool
- Choose one local PDF up to 20 MB and 20 pages.
- Convert the pages in your browser at 96 DPI; each page image and extracted text are embedded in the HTML.
- Download and open the standalone HTML file. The PDF is never uploaded.
A portable, self-contained HTML document
Every PDF page is rendered as an embedded PNG image, preserving its visible layout. Existing PDF text is included separately in a selectable transcript below that page. Images and text are embedded into one HTML download; no remote images, scripts, fonts, or network requests are added.
Local processing and limits
The selected PDF stays in this browser. One PDF may be up to 20 MB and 20 pages. Pages are rendered at 96 DPI, with a maximum of 4 megapixels per page, 30 megapixels total, and 100 MB of HTML output.
What the HTML does not preserve
This is a visual page archive, not a semantic or editable reconstruction of the PDF. Vector artwork becomes PNG images; text is a separate transcript and does not align over the page. Scanned pages remain visible as images but receive no OCR transcript. Links, forms, annotations, signatures, document structure, and metadata are not carried over.
GOOD TO KNOW
Common questions
Does the HTML keep the original page appearance?
It embeds a rendered PNG of every page at 96 DPI. This preserves visible page content as an image, not editable PDF vectors or objects.
Can I select or search the text?
Embedded PDF text is also included as a separate text transcript for each page. It may not match the page layout or reading order. Scanned image-only pages are not OCRed.
Is the downloaded file self-contained?
Yes. Page images and transcripts are embedded directly in one HTML file; it does not need external assets or scripts.
Is my PDF uploaded?
No. PDF rendering, text extraction, and HTML creation all happen in your browser.