qli917/html_skill

Extract readable main content from messy HTML and URLs for AI agents, RAG pipelines, summarization, and scraping cleanup.

2Stars on the repository
1Mods indexed here, across every type
2mo agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

extract-html-main

01

qli917/html_skill

Skill Claude CodeCodex

Extract readable main body content from arbitrary HTML pages, local HTML files, saved browser pages, or URLs where the article/body structure is unknown or inconsistent. Use when Codex needs to remove navigation, ads, boilerplate, comments, sidebars, scripts, hidden text, duplicated menus, or layout chrome and return…

2 2mo ago A 93 tokens original MIT