defuddle

Extract clean markdown from web pages by removing ads and navigation.

Updated Apr 23, 2026
One-click install
npx skills add https://github.com/xiang2007/obsidian-vault --skill defuddle-xiang2007
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: defuddle
Source: https://github.com/xiang2007/obsidian-vault/tree/main/skills/defuddle
Command: npx skills add https://github.com/xiang2007/obsidian-vault --skill defuddle-xiang2007

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill streamlines the process of ingesting web content into a wiki by removing unnecessary elements like ads and navigation, resulting in a more concise and token-efficient markdown.

Core Features & Use Cases

  • Web Page Cleanup: Strip ads, navigation, headers, footers, and boilerplate from web pages.
  • Token Efficiency: Reduces token usage by 40-60% on typical web articles.
  • Use Case: Before adding a new article from a URL to your Obsidian wiki, use defuddle to ensure the content is clean and efficient.

Quick Start

Use the defuddle skill to clean the content from the URL 'https://example.com/article'.

Frequently Asked Questions about defuddle

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I clean web content into markdown for Obsidian?

To clean web content into markdown for Obsidian, you need to extract meaningful text while stripping ads and navigation. This skill processes web pages to produce concise, token-efficient markdown suitable for your personal knowledge management system.

How do I reduce token usage when ingesting web articles?

You can reduce token usage during web article ingestion by removing unnecessary boilerplate and ads. Cleaning web content before processing it as markdown typically decreases token consumption by 40 to 60 percent, maintaining a lean content library.

Does cleaning web page content remove ads and navigation elements?

Yes, cleaning web page content removes ads, navigation, headers, footers, and other boilerplate. This extraction process ensures only meaningful article text remains, resulting in more efficient markdown output for your knowledge base.

Can I use this to extract meaningful content before adding URLs to my wiki?

Yes, you can use this to extract meaningful content before adding URLs to your wiki. It processes raw web pages to remove unnecessary elements, ensuring the content you ingest into your Obsidian wiki is clean and efficient.

What is the best way to strip boilerplate from web pages for personal knowledge management?

The best way to strip boilerplate from web pages for personal knowledge management is to use a content cleaning process that removes ads and navigation. This yields concise markdown, optimizing token efficiency for your Obsidian wiki.