defuddle
kepano/obsidian-skills
Extract clean Markdown from HTML pages, removing ads and clutter to reduce token usage.
What is defuddle?
Defuddle CLI parses web pages and extracts readable Markdown content, stripping navigation, ads, and other noise. Use it when you need clean text from standard web pages with minimal processing overhead.
- Parse URLs and extract clean Markdown output with `--md` flag
- Remove navigation, ads, and page clutter automatically
- Save extracted content directly to Markdown files with `-o` option
- Extract specific metadata properties (title, description, domain) from pages
- Output in multiple formats: Markdown, JSON, or raw HTML
How to install defuddle
npx skills add https://github.com/kepano/obsidian-skills --skill defuddle- Node.js and npm installed
- Install globally with: `npm install -g defuddle`
How to use defuddle
- 1.Install Defuddle globally: `npm install -g defuddle`
- 2.Parse a URL to Markdown: `defuddle parse <url> --md`
- 3.Save output to file: `defuddle parse <url> --md -o content.md`
- 4.Extract specific metadata: `defuddle parse <url> -p title` (or description, domain)
- 5.Choose output format: use `--md` for Markdown, `--json` for structured data, or omit flag for HTML
Use cases
- Convert web articles to clean Markdown for note-taking or documentation
- Extract article content while filtering out sidebars and ads before processing
- Batch download and convert multiple web pages to Markdown files
- Pull specific metadata like titles and descriptions from web pages
- Reduce token usage by cleaning HTML before feeding to language models
- Knowledge workers managing web research and notes
- Developers building content pipelines or web scrapers
- AI agents processing web content for analysis or summarization
- Technical writers converting online resources to documentation
defuddle FAQ
Use Defuddle for standard web pages where you want clean, readable content with ads and navigation removed. It's more efficient for typical web scraping and reduces token usage.
Defuddle supports Markdown (--md), JSON (--json with both HTML and markdown), raw HTML (no flag), and specific metadata properties (-p flag).
Yes, use the `-o` flag followed by a filename: `defuddle parse <url> --md -o content.md`
Use the `-p` flag with the property name: `defuddle parse <url> -p title` or `defuddle parse <url> -p description`
Full instructions (SKILL.md)
Source of truth, from kepano/obsidian-skills.
name: defuddle description: Extract clean Markdown from HTML pages with Defuddle CLI.
Defuddle
Use Defuddle CLI to extract clean readable content from web pages. Prefer over WebFetch for standard web pages — it removes navigation, ads, and clutter, reducing token usage.
If not installed: npm install -g defuddle
Usage
Always use --md for markdown output:
defuddle parse <url> --md
Save to file:
defuddle parse <url> --md -o content.md
Extract specific metadata:
defuddle parse <url> -p title
defuddle parse <url> -p description
defuddle parse <url> -p domain
Output formats
| Flag | Format |
|---|---|
--md | Markdown (default choice) |
--json | JSON with both HTML and markdown |
| (none) | HTML |
-p <name> | Specific metadata property |
Related skills
More from kepano/obsidian-skills and the wider catalog.

json-canvas
Create and edit JSON Canvas files (.canvas) with nodes, edges, groups, and connections for visual diagrams in Obsidian.

knap
Render Markdown from templates and structured data using Knap CLI.

obsidian-bases
Create and edit Obsidian Bases (.base files) with views, filters, formulas, and summaries.

obsidian-cli
Interact with Obsidian vaults from the command line—read, create, search, manage notes, and develop plugins.

obsidian-markdown
Create and edit Obsidian Flavored Markdown with wikilinks, embeds, callouts, and properties.

text-watermark-cleaner-zh-tw
Remove invisible Unicode, zero-width characters, and statistical AI watermarks from Traditional Chinese text while preserving formatting and meaning.