HTML Tag Remover

HTML markup turned into readable plain text

Preset
Read only

What is HTML Tag Remover?#

HTML Tag Remover turns markup from a web page or an email into plain text. Tags are removed and character references are resolved, so the result is readable text with headings and paragraphs still separated.

How to Use#

  1. Paste HTML into the input box.
  2. Adjust the result with a preset or the individual options when the default is not what you need.
  3. Copy the result and use it where you need it.

Options#

  • Block line breaks: puts line breaks around paragraphs, headings, list items, and table rows. List items get a bullet. Turn it off to join everything into running text.
  • Tidy whitespace: collapses consecutive blank lines into one and trims each line. Turn it off to keep the original spacing.
  • Remove scripts and styles: drops script, style, and noscript content along with the tags. Leave it on unless you need to see that text.
  • Show link URLs: writes links as text (URL). A link with no text shows its URL alone.

Three presets set these at once. Standard keeps structure without URLs, Keep links adds the URLs, and Unformatted turns shaping off while still dropping scripts and styles.

Use Cases#

  • Pulling article text out of a saved web page for quoting or summarizing.
  • Cleaning markup out of an HTML email before replying or filing it.
  • Turning scraped pages or API responses into text for analysis.
  • Lifting the values in a table into a spreadsheet.

FAQ#

Notes#

  • Content in the document head, such as the page title, is not included. Only the body text comes out.
  • HTML comments are always removed.
  • Formatter

Last updated:

Related tools