WrenPDF

PDF → HTML

Turn PDF text into HTML, Markdown or plain text while preserving heading, paragraph and list structure.

What it does

Copying plain text out of a PDF loses headings and lists the moment you paste it into a website, blog or notes app. This tool infers headings from font sizes and paragraphs from line spacing, and gives you HTML, Markdown or plain text.

Preview the output before downloading or copy it to the clipboard. Pick a page range to convert only the part you need.

How to use it

  1. Drop your PDF file into the box.
  2. Choose the output format (HTML, Markdown, text) and the page range.
  3. Convert it, preview it, then download or copy it.

When is it useful?

  • Moving PDF articles onto a blog or website
  • Importing documents into Notion or Obsidian as Markdown
  • Adding PDF text to an email newsletter
  • Turning PDF content into an accessible web page

Tips

  • Markdown gives the cleanest result in note-taking apps.
  • For table-heavy content, PDF → Excel is the better choice.

Frequently asked questions

Can text be extracted from scanned (photo-like) PDFs?
No. That requires OCR; for now only PDFs that have a text layer are supported.
Are tables and multi-column pages preserved?
The output uses a flowing layout; tables and columns are turned into rows in reading order.
Where does the text extraction run?
The pdf.js parsing happens inside a Web Worker; the interface stays responsive even with large PDFs.
Are images included in the output?
No, the output focuses on text. Use PDF Extractor to get the images separately.

Related tools