Skip to content
PDFToolz

PDF to HTML Converter

Convert a PDF to a web page. Get clean HTML with real headings and tables, or an exact copy of each page with selectable text.

Select a PDF file

or drop it anywhere on this page, or paste

Opens in your browser. Nothing is uploaded.

PDF to HTML Converter, free and private

Convert PDF to HTML for free and pick the kind of web page you need. Clean HTML rebuilds the document as real h1 to h6 headings, paragraphs, lists and tables that reflow on a phone. Exact layout places every line of text where it sits on the PDF page, over a picture of the page, so the result looks like the original.

Both modes save one self-contained .html file with images stored inside it, or a ZIP with the images as separate files. The PDF is converted in your browser and never uploaded, and a live preview of the HTML appears before you download.

How to convert a PDF to HTML

  1. 1

    Add your PDF

    Press Select a PDF file or drop the PDF onto the page. The tool reads one document at a time and shows a preview as soon as the first pages are read.

  2. 2

    Pick Clean HTML or Exact layout

    Clean HTML suits web pages, blogs, email templates and content management systems. Exact layout suits archives and pages that must look like the printed document.

  3. 3

    Set images and the page background

    In Clean HTML, choose whether to include the images. In Exact layout, choose Page picture, Graphics only or None. Then choose one .html file or a ZIP, and all pages or a range like 1-4, 8.

  4. 4

    Convert PDF to HTML and download

    Press Convert to HTML, then Download HTML or Download ZIP. Open the file in any browser or paste its body into your site.

Why use PDFToolz for this

Clean, semantic HTML

Headings get h1 to h6 by font size, wrapped lines are joined into p elements, bullets and numbers become nested ul and ol lists, and aligned columns become a table with a thead row. The file carries a small style block and no inline styles, so your own CSS takes over when you paste the markup into a page.

Exact layout with selectable text

Each text run is an absolutely positioned span, measured in points, on a page the size of the PDF page. Rotated text stays rotated, and a short script stretches each run to its width in the PDF, so lines end where they did in the original.

Three page backgrounds

Page picture draws the whole page at 150 DPI and lays invisible text on top, so it looks exactly like the PDF and you can still select and search the text. Graphics only draws the shapes and images without the text, and the real HTML text sits on top. None keeps the text alone for the smallest file.

One file or a ZIP

One .html file stores images as base64 data URIs, so it opens anywhere and you can send it as a single attachment. The ZIP keeps the images as JPG or PNG files beside the HTML, which is about a quarter smaller and easier to upload to a web server.

Links that still work

Web and email links in the PDF become a elements in both modes. Other link types, such as scripts or file links, are dropped.

Free, private and without Acrobat

No account, watermark or page limit, and no Adobe Acrobat needed. It runs in Chrome, Edge, Firefox and Safari on Windows, Mac, Linux, iPhone, iPad and Android. Your PDF stays on your device.

Choosing between clean HTML and exact layout

Pick Clean HTML when the content matters more than the look. It reads well on a phone, works with screen readers and search engines, and drops into WordPress, Webflow or an email builder. The PDF's fonts, colours and positions are not kept.

Pick Exact layout when the page has to look like the PDF, such as a form, a brochure or a scanned archive page. Each page keeps its size, so the HTML scrolls sideways on narrow screens. It is a faithful copy, not a page you would edit by hand.

How to convert PDF to HTML for a website

Convert with Clean HTML and save as a ZIP. Upload the images folder to your site, then copy everything between the body tags of the .html file into your page or CMS. Headings, lists and tables keep their meaning, and your site's styles apply to them.

Repeated headers, footers and page numbers are removed by default, so they do not show up in the middle of the article. Turn on Page markers to keep a dashed rule and a <!-- Page N --> comment at each page break.

What the exact layout mode keeps

Positions, font sizes, bold, italic and rotation come from the PDF. The browser shows the text in its own sans-serif, serif or monospace font, because the PDF's embedded fonts are not copied. The width of every run is matched to the PDF, so small differences in letter shapes are the only visible change.

  • Page picture: a JPEG of each page at 150 DPI, capped at 4 million pixels for very large pages. Colours, images and fonts look as in the PDF.
  • Graphics only: the same picture with the text left out. Text is shown in black, because PDF text colours are not read.
  • None: only the text and links, with no picture.

Scanned PDFs and other limits

A scanned PDF has no text layer. Clean HTML and Exact layout with no background stop and point you to OCR PDF. Exact layout with Page picture still converts a scan, but the result is pictures of the pages with no text to select. Run OCR PDF first to get both.

Layouts with three or more text columns are read across the page in Clean HTML. Tables with merged cells can come out with extra rows. Form field values and comments are not converted. A PDF that needs a password to open has to go through Unlock PDF first. Page pictures make files large, so a 100-page PDF with Page picture can reach tens of megabytes.

Frequently asked questions

Related tools and guides