html to markdown skills

30 tagged html to markdown, measured the same way as everything else here.

Browse within: crawler 20docker 16data-extraction 15markdown 7browser-automation 6scraper 5

firecrawl/firecrawl

Skill Claude CodeCodex

Get Firecrawl credentials and SDK setup into a project. Use when an application needs FIRECRAWLAPIKEY, when an agent should add Firecrawl to .env, when the user wants to authenticate Firecrawl for app code, or when choosing the first SDK and docs for a new Firecrawl integration. This skill includes its own browser…

175k +394 today A 89 tokens AGPL-3.0

firecrawl/firecrawl

Skill Claude CodeCodex

Integrate Firecrawl /scrape into product code for single-page extraction. Use when an app already has a URL and needs markdown, HTML, links, screenshots, metadata, or structured page output. Prefer this skill over broader crawl patterns when the feature is page-level.

175k +394 today A 62 tokens AGPL-3.0

firecrawl-build

03

firecrawl/firecrawl

Skill Claude CodeCodex

Integrate Firecrawl into application code whenever a product, agent, or workflow needs web data inside the app — web search, live search results, page scraping, structured extraction, or browser interaction. Use when building any feature that needs data from the web in code, even if the user does not mention Firecrawl…

175k +394 today A 149 tokens AGPL-3.0

crw

04

us/crw

Skill Claude CodeCodex

Scrape, crawl, map, and search the web using fastCRW's native /v1 API. Use when the user needs web page content, site-wide extraction, URL discovery, or web search results. Single binary, 14 MB RAM; /v2 exists separately for Firecrawl migration.

882 2d ago A 64 tokens AGPL-3.0

crw-best-practices

05

us/crw

Skill Claude CodeCodex

Reference skill for building production-ready crw integrations. Covers verb selection, call surfaces (CLI/MCP/REST), post-filtering strategies, context-window hygiene, Hybrid RAG patterns, common pitfalls, and crw-specific operational considerations (search backend limits, renderer pool, proxy rotation). Load this…

882 2d ago A 99 tokens AGPL-3.0

crw-dynamic-search

06

us/crw

Skill Claude CodeCodex

Programmatic web search and scrape with context isolation. Use for any research task where you need to search the web, filter results, and extract specific information — without flooding your context window with raw HTML and boilerplate. This is the single biggest token-saver in the crw skill set. Triggered by "search…

882 2d ago A 143 tokens AGPL-3.0

pullmd

07

AeternaLabsHQ/pullmd

Skill Claude CodeCodex

Read any web page, document, or YouTube video as clean Markdown using PullMD. Use this skill whenever you need to fetch, read, extract, or summarize content from a URL — web articles, Reddit threads, PDF/Word/PowerPoint/Excel/EPUB documents, or YouTube transcripts. This includes when the user says 'read this page'…

480 5d ago B 159 tokens AGPL-3.0

agent-fetch

08

teng-lin/agent-fetch

Skill Claude CodeCodex

Fetch and extract full article content from URLs. Returns complete text with structure (headings, links, lists) instead of summaries. Multiple extraction strategies, browser impersonation, cookies, crawling, custom selectors, 200-700ms.

307 5mo ago A 50 tokens original MIT

cnkang/nginx-markdown-for-agents

Skill Claude CodeCodex

Route and validate harness maintenance work for nginx-markdown-for-agents. Use when changing AGENTS.md, docs/harness, tools/harness, Makefile, or CI harness wiring. Also use it when you need spec resolution, risk-pack routing, or phased verification commands.

23 2d ago A 64 tokens original BSD-2-Clause

agentic-search

10

appautomaton/webmaton

Skill Claude CodeCodex

Grok-primary deep research skill for source-backed web work. Use when the task needs current web research, grounded citations, supplementary Tavily/Firecrawl source discovery, high-fidelity page-to-Markdown fetch, Tavily site mapping, verbatim quote extraction, source reranking, or reusable multi-step research…

19 7d ago A 103 tokens original MIT

nodriver-browser

11

appautomaton/webmaton

Skill Claude CodeCodex

Persistent Chrome/Chromium browser automation skill built on nodriver. Use when a page needs JavaScript rendering, authorized login/session continuity, clicking or typing, DOM snapshots with stable refs, screenshots, or multi-step look-think-act flows that ordinary WebFetch/search cannot complete. Auto-starts a…

19 7d ago A 106 tokens original MIT

playwright-skill

12

appautomaton/webmaton

Skill Claude CodeCodex

Complete browser automation with Playwright. Auto-detects dev servers, writes clean test scripts to $TMPDIR (or /tmp). Test pages, fill forms, take screenshots, check responsive design, validate UX, test login flows, check links, automate any browser task. Use when user wants to test websites, automate browser…

19 7d ago A 83 tokens original MIT

html-to-markdown

13

appautomaton/markmaton

Skill Claude CodeCodex

Convert a URL or HTML into clean Markdown with metadata using markmaton. Handles browser capture for JS-heavy pages and deterministic HTML-to-Markdown conversion in one skill.

5 18d ago A 38 tokens original MIT