HTML to Markdown

Turn saved web pages or exported HTML into readable Markdown — headings, lists, tables, links and code blocks intact, markup and scripts gone. Free in your browser — no sign-up, no watermarks.

Drop your HTML files here

or click to browse — Select multiple files for batch conversion

htmlhtm

⚡ Instant — conversion starts the moment you add files. No sign-up, no watermarks. Why FileHugger

🔍 How your file is processed
Runs
In your browser on this page
Engine
Our own Markdown parser, OOXML writer and ZIP reader; HTML is parsed by the browser’s own parser
Max input
512 MB per file
Your file
Processed in this browser tab — the file is not uploaded. Converted results are kept in this browser's local storage for up to two hours so their matching download page can show them; expired results are removed on the next site visit. Nothing is stored on our servers.
Good to know
Structure converts, presentation does not. Headings, paragraphs, lists, tables, bold, italic, code, quotes and links all survive; fonts, colours, page layout, headers and footers, footnotes, comments and tracked changes do not. Images inside a Word document are not carried into Markdown, which has nowhere to store them.

Full security & file-privacy details →

Related converters

How to turn a web page into Markdown

  1. Click Choose files (or drag & drop) and select one or more HTML files. You can also paste an image from your clipboard.
  2. No settings needed for this conversion — the defaults handle everything.
  3. Press Convert to Markdown. Each file is converted with a live progress indicator.
  4. Open the download page to save your files individually or grab everything as a single ZIP.

Why convert HTML to Markdown?

HTML is what content management systems, help desks, newsletter tools and saved web pages export. Markdown is what documentation, notes and static sites want. Moving between them by hand means stripping tags one at a time; a regex approach breaks on the first unclosed tag. This uses the browser’s own HTML parser, so it copes with real-world markup the way a browser does, and keeps the structure that matters while dropping the styling, scripts and layout wrappers that do not.

Advertisement

About the formats

What is an HTML file?

HTML is the markup language of the web — the format every browser renders. Generating clean HTML from Markdown gives you ready-to-publish markup for websites, newsletters and CMS content without writing tags by hand.

What is a Markdown file?

Markdown is a lightweight way to write formatted text with plain characters: # for headings, ** for bold, - for lists. READMEs, documentation, notes apps and static site generators all speak Markdown. Converting Markdown to HTML produces the exact markup browsers render, ready to paste into a website, email template or CMS.

HTML vs Markdown at a glance

HTMLMarkdown
TypeText markupText markup
CompressionTextText
Typical useWeb pages, newsletters, CMS contentDocumentation, READMEs, notes

Real HTML is not well-formed, and that decides the approach

HTML is what content systems export: help centres, newsletter tools, blog platforms, and every web page anyone has ever saved. Markdown is what documentation and notes want. The gap between them gets crossed by hand more often than it should.

The tempting way to do this is with regular expressions over the markup, and it falls apart immediately, because published HTML is full of unclosed tags, attributes without quotes and nesting no pattern survives. This uses the browser’s own HTML parser instead — the same one that renders the page — so malformed markup is handled exactly the way a browser handles it, and what gets walked afterwards is a real tree.

The page, not the furniture

A saved web page is mostly not the article: it is navigation, a sidebar, a cookie notice and a footer. When the page marks its content with an article or main element — most published pages do — only that is converted. Otherwise the whole body is, and you may want to trim the ends.

What survives

Headings at their levels, paragraphs, bulleted and numbered lists including nested ones, tables with their header row, block quotes, preformatted blocks as fenced code, images as Markdown images, and links with their addresses. Bold, italic, strikethrough and inline code come through as their Markdown equivalents. Characters that mean something in Markdown are escaped, so a price of *99 in the source does not turn the rest of the paragraph italic.

What is removed, and why

Scripts, stylesheets, frames and embedded objects are stripped before anything is read. Parsing happens in an inert document, so nothing in the input loads, runs or fetches while it is being converted — which matters, because the usual reason to convert HTML is that it came from somewhere else. Links that point at javascript: become plain text rather than links.

Frequently asked questions

Is this HTML to Markdown tool free?

Yes — completely free, with no sign-up, no watermarks and no daily quota. Processing happens in this browser, so the practical limit is the memory available on your device.

Can I convert multiple HTML files at once?

Yes. Select or drop as many files as you like — each one is converted in sequence with its own progress status, and you can download the results individually or all together as a ZIP archive.

Does it convert the whole page or just the content?

If the page marks its main content with an article or main element — most published pages do — that part is converted and the navigation, sidebars and footer are left out. Otherwise the whole body is converted. Scripts, styles and embedded frames are removed before anything is read.

Advertisement