content extraction skills

34 tagged content extraction, measured the same way as everything else here.

Browse within: anycrawl 16claude-code-skills 16ai-integration 8github-explorer 6multi-source-search 6openclaw 6search 6

agent-fetch

01

teng-lin/agent-fetch

Skill Claude CodeCodex

Fetch and extract full article content from URLs. Returns complete text with structure (headings, links, lists) instead of summaries. Multiple extraction strategies, browser impersonation, cookies, crawling, custom selectors, 200-700ms.

305 5mo ago A 50 tokens original MIT

you-discover

02

youdotcom-oss/agent-skills

Skill Claude CodeCodex

Route You.com integration planning through the you-discover MCP tool, Docs MCP, and direct API options.

64 7d ago A 25 tokens original MIT

you-finance

03

youdotcom-oss/agent-skills

Skill Claude CodeCodex

Route finance questions to an existing local script, a new You.com Finance Research API call, or an MCP payment-aware fallback.

64 7d ago A 29 tokens original MIT

you-web

04

youdotcom-oss/agent-skills

Skill Claude CodeCodex

Use You.com MCP tools for current web search, URL content extraction, cited web synthesis, and x402-aware web access.

64 7d ago A 28 tokens original MIT

content-extract

05

blessonism/openclaw-skills

Skill Claude CodeCodex

Robust URL-to-Markdown extraction for OpenClaw workflows. Use when the user wants to "extract/summarize/convert a webpage to markdown" (especially WeChat mp.weixin.qq.com) and webfetch/browser is blocked or messy. Uses a cheap probe via webfetch first, then falls back to the official MinerU API (via the local…

55 5mo ago A 92 tokens original MIT

dependency-tracker

06

blessonism/openclaw-skills

Skill Claude CodeCodex

Track and check updates for all OpenClaw dependencies: managed skills (GitHub/ClewHub), bundled skills, workspace skills, npm packages, pip packages, and CLI tools. Use when user asks "check for updates", "dependency status", "are my skills up to date", "什么需要更新", "检查依赖", "检查更新", or wants a dependency health report.…

55 5mo ago A 98 tokens original MIT

search-layer

07

blessonism/openclaw-skills

Skill Claude CodeCodex

DEFAULT search tool for ALL search/lookup needs. Multi-source search and deduplication layer with intent-aware scoring. Integrates Brave Search (websearch), Exa, Tavily, and Grok to provide high-coverage, high-quality results. Automatically classifies query intent and adjusts search strategy, scoring weights, and…

55 5mo ago A 116 tokens original MIT

searchts

08

capad-xyz/searchts

Skill Claude CodeCodex

MUST USE for web research, lookup, or any shared URL/link (Twitter/X, Reddit, YouTube, GitHub, LinkedIn, or any page). Read with searchts read — unlocker for 403/429/bot-walls and JS. Search with searchts search. Transcribe with searchts transcribe. Grab assets/palette with searchts grab or searchts get. Prefer…

2 2d ago A 110 tokens original MIT

citedy/skills

Skill Claude CodeCodex

Parallel multi-agent code review using Agent Teams with 4 specialized reviewers. Spawns a coordinated team of security, performance, test coverage, and code quality agents. Teammates can share findings with each other for cross-domain insights. Produces unified report with severity-ranked findings saved to /output/.…

2 3mo ago A 204 tokens original MIT

prompt-analyzer

10

citedy/skills

Skill Claude CodeCodex

Analyze prompts for constraint complexity, audit failure risks, and generate optimized rewrites for Claude and GPT. Based on "How LLMs Follow Instructions" (Rocchetti & Ferrara, 2026) constraint taxonomy research. Use when reviewing prompt files, optimizing prompt bases, or auditing instruction quality. Trigger…

2 3mo ago A 95 tokens original MIT

schema-markup

11

citedy/skills

Skill Claude CodeCodex

When the user wants to add, fix, or optimize schema markup and structured data on their site. Also use when the user mentions "schema markup," "structured data," "JSON-LD," "rich snippets," "schema.org," "FAQ schema," "product schema," "review schema," or "breadcrumb schema." For broader SEO issues, see seo-audit.

2 3mo ago A 78 tokens original MIT

extract-html-main

12

qli917/html_skill

Skill Claude CodeCodex

Extract readable main body content from arbitrary HTML pages, local HTML files, saved browser pages, or URLs where the article/body structure is unknown or inconsistent. Use when Codex needs to remove navigation, ads, boilerplate, comments, sidebars, scripts, hidden text, duplicated menus, or layout chrome and return…

2 2mo ago A 93 tokens original MIT