document processing skills

167 tagged document processing, measured the same way as everything else here.

Browse within: ai-assistant 52data-analysis 52docx 23claude-ai 20python 20knowledge-graph 17ETL 16agent-skill 16claude-skill 16lightrag 16ocr 16anki-flashcards 9document 9markdown-converter 9

handoff

25

maaarcooo/agent-skills

Skill Claude CodeCodex

Create a structured handoff document for seamless continuity across Claude conversations. Use when the user explicitly asks to create a handoff, handover, or session transfer document, or signals they want to stop now and continue this work in a future conversation. Trigger phrases include: "create a handoff"…

10 3d ago A 181 tokens

learn

26

maaarcooo/agent-skills

Skill Claude CodeCodex

Use this skill when the user wants intellectual understanding — learning how or why something works, not getting a task done or soliciting Claude's judgment. Trigger for: Explicit learning requests: teach, explain, ELI5, walk me through, quiz me, flashcards, "I'm rusty on"; definitions ("what is X") Terse concept…

10 3d ago A 219 tokens

7

27

ranbot-ai/awesome-skills

Skill Claude CodeCodex

Security audit, hardening, threat modeling (STRIDE/PASTA), Red/Blue Team, OWASP checks, code review, incident response, and infrastructure security for any project.

6 2d ago A 39 tokens

ranbot-ai/awesome-skills

Skill Claude CodeCodex

AI-powered presentation generation via the 2slides API — create slides from text, match a reference image style, summarize documents into decks, add AI voice narration, and export pages/audio. Use f.

6 2d ago A 45 tokens

che-word-mcp

29

PsychQuant/che-word-mcp

Skill Claude CodeCodex

A Swift-native MCP server for Microsoft Word (.docx) document manipulation. Provides 83 tools for reading, writing, and modifying Word documents without requiring Microsoft Word installation.

6 2d ago A 0 tokens original MIT

extracto-cli

30

codelined-ag/Extracto

Skill Claude CodeCodex

Use when the user wants to extract text from PDFs or images, manage OCR jobs, or work with output presets via the local Extracto OCR webapp. Talks HTTP to a running Extracto instance using the bundled extracto CLI.

5 3d ago A 52 tokens original MIT

book-skills-creator

31

Tire-C/book-skills-creator

Skill Claude CodeCodex

Create modular, source-grounded Agent Skill packs from explicitly selected books or documents. Use when a user wants atomic skills, combo workflows, a router, references, a skill map, and validation derived from a specific source.

4 1mo ago A 50 tokens original MIT

deepread-api

34

deepread-tech/skills

Skill Claude CodeCodex

Full DeepRead API reference. All endpoints, auth, request/response formats, blueprints, webhooks, error handling, and code examples for OCR, structured extraction, form filling, and PII redaction.

4 1mo ago A 47 tokens original MIT

deepread-form-fill

35

deepread-tech/skills

Skill Claude CodeCodex

AI-powered PDF form filling via DeepRead. Upload any PDF form + your data as JSON — AI detects fields, maps data semantically, fills the form with quality checks, returns a completed PDF. Works with scanned, non-editable forms — no AcroForm required.

4 1mo ago A 60 tokens original MIT

deepread-pii

36

deepread-tech/skills

Skill Claude CodeCodex

Redact PII from documents before sharing or sending to LLMs. 14 PII types (names, SSN, credit cards, medical records, etc.) detected with context-aware AI — not regex. Knows patient vs. doctor, personal vs. institutional. Black bar redaction on PDFs, scanned images, and text files. Free tier: 2,000 pages/month.

4 1mo ago A 84 tokens original MIT

rlm-chunking

37

zircote-plugins/rlm-rs-plugin

Skill Claude CodeCodex

This skill should be used when the user asks about "chunking strategies", "how to chunk a file", "semantic vs fixed chunking", "chunk size selection", "overlap settings", or needs guidance on selecting the appropriate chunking approach for RLM processing.

3 1mo ago A 60 tokens

rlm

38

zircote-plugins/rlm-rs-plugin

Skill Claude CodeCodex

This skill should be used when the user asks to "process a large file", "analyze document exceeding context", "use RLM", "recursive language model workflow", "chunk and analyze a file", "handle long context", or needs to work with documents too large to fit in the context window. Orchestrates the full RLM loop using…

3 1mo ago A 80 tokens

gootf/corpus-knowledge-engineering

Skill Claude CodeCodex

Turn document corpora (books, merged TXT, PDF/EPUB, OCR text) into structured agent knowledge: normalize → deterministic source segmentation → chapter maps with provenance → book skills → evaluation gates. Use when the user wants to convert books/materials into AI skills, build a knowledge base from a corpus, evaluate…

2 2d ago A 89 tokens original MIT

doc-compression

41

fedeclavero/doc-compression-skill

Skill Claude CodeCodex

Compresión textual por eliminación / text compression by deletion. Acorta cualquier texto eliminando redundancias, repeticiones y prosa decorativa, preservando estructura, terminología, definiciones, citas, datos factuales y secuencia argumentativa. NO reescribe ni parafrasea. Usar cuando el usuario pida comprimir…

2 1mo ago A 193 tokens original MIT

pdf-mcp

42

sandraschi/pdf-mcp

Skill Claude CodeCodex

Tool-awareness for the pdf-mcp PDF intelligence MCP server. Load when working with pdf-mcp tools or the pdf-mcp webapp.

1 13d ago A 32 tokens original MIT

pdf-expert

43

sandraschi/pdf-mcp

Skill Claude CodeCodex

Skill "pdf-expert" from sandraschi/pdf-mcp, covering pdf-mcp skill, tool categories, pdfextract — extract content from pdfs, pdfmanipulate — modify pdf structure and pdfannotate — add markup and annotations.

1 13d ago A 0 tokens original MIT

skills

44

adrianwedd/rlm-mcp

Skill Claude CodeCodex

Activate RLM processing when ANY of the following apply.

1 3mo ago A 0 tokens

pdf-asset-extractor

48

u9401066/asset-aware-mcp

Skill Claude CodeCodex

MCP tools for PDF ingestion → extract figures, tables, sections → build knowledge graph. Transforms PDF into queryable assets (images, tables, text) with cross-document RAG. Triggers: PDF, ingest, extract, 圖片, 表格, figure, table, manifest, knowledge graph, 知識圖譜, RAG, 文獻分析.

0 14d ago A 84 tokens copy · 100% Apache-2.0