# Extract clean article Markdown from web pages with Defuddle

> Use Defuddle when an agent needs clean, metadata-rich article text or Markdown from noisy web pages before summarizing, indexing, or archiving them.

- Skill: `agentskillexchange/extract-clean-article-markdown-from-web-pages-with-defuddle` (Agent Skill)
- Install (CLI): `npx skillmds@latest add agentskillexchange/extract-clean-article-markdown-from-web-pages-with-defuddle`
- Raw SKILL.md: https://api.skillmd.com/api/skills/agentskillexchange/extract-clean-article-markdown-from-web-pages-with-defuddle/raw
- Safety review: pending (external: skill-scanner PASS, skillspector PASS)
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: AI & ML
- Author: agentskillexchange (https://skillmd.com/u/agentskillexchange)
- Updated: 2026-09-08
- Page: https://skillmd.com/skills/agentskillexchange/extract-clean-article-markdown-from-web-pages-with-defuddle

---


# Extract clean article Markdown from web pages with Defuddle

Use Defuddle when an agent needs clean, metadata-rich article text or Markdown from noisy web pages before summarizing, indexing, or archiving them.

## Prerequisites

Node.js, npx or npm, defuddle CLI

## Installation

Use the upstream install or setup path that matches your environment:
- npx defuddle parse page.html
- npx defuddle parse https://example.com/article
- npx defuddle parse page.html --markdown
- npx defuddle parse page.html --json

Requirements and caveats from upstream:
- ### Node.js
- defuddle/node accepts a DOM Document from any implementation (JSDOM, linkedom, happy-dom, etc.).
- import { Defuddle } from 'defuddle/node';

Basic usage or getting-started notes:
- Defuddle takes a URL or HTML, finds the main content, and returns cleaned HTML or Markdown. Defuddle was created for the browser extension [Obsidian Web Clipper](https://github.com/obsidianmd/obsidian-clipper), but it...
- ### Browser
- javascript

- Source: https://github.com/kepano/defuddle
- Extracted from upstream docs: https://raw.githubusercontent.com/kepano/defuddle/HEAD/README.md

## Documentation

- https://defuddle.md

## Source

- [Agent Skill Exchange](https://agentskillexchange.com/skills/extract-clean-article-markdown-from-web-pages-with-defuddle/)

