document parser skills

10 tagged document parser, measured the same way as everything else here.

autorag-setup

01

Marker-Inc-Korea/AutoRAG

Skill Claude CodeCodex

Configure AutoRAG for first use or repair its single-agent model, approved document roots, retrieval indexes, datasource skills, and health checks without exposing credentials.

5.1k 3d ago A 36 tokens

autorag

02

Marker-Inc-Korea/AutoRAG

Skill Claude CodeCodex

Use an already configured AutoRAG librarian agent to search, summarize, compare, and answer questions from local document collections. Use autorag-setup for configuration or indexing changes.

5.1k 3d ago A 38 tokens

gongmunseo

03

chrisryugj/kordoc

Skill Claude CodeCodex

A tool for turning supplied content into Korean government-style official documents in HWPX format. HWPX is a Korean word-processing document format.

1.8k 3d ago A 286 tokens original MIT

kordoc

04

chrisryugj/kordoc

Skill Claude CodeCodex

Use this skill whenever the user wants to read, create, fill, edit, compare, validate, or preview Korean Hangul/official documents — .hwp (HWP 3.x/5.x), .hwpx, .hml (HWPML) — or convert Korean-office PDF/DOCX/XLS/XLSX to Markdown. Triggers include any mention of 'hwp', 'hwpx', 'hml', '한글 문서', '아래한글', '한컴', '공문서'…

1.8k 3d ago A 253 tokens original MIT

deckprobe

05

deckflow/deckprobe

Skill Claude CodeCodex

Inspect PDF, Microsoft Office (docx/xlsx/pptx and legacy doc/xls/ppt), and Apple iWork (key/numbers/pages) files without opening or rendering them. Use for page and slide and sheet counts, title/author/created metadata, encryption and macro and digital-signature and JavaScript risk signals, sheet and table structure…

82 7d ago C 130 tokens original MIT

all2md

06

thomas-villani/all2md

Skill Claude CodeCodex

Convert, read, generate, search, and compare documents with the all2md CLI and Python library. Use whenever a task involves reading or extracting text/tables from a document (PDF, Word, PowerPoint, Excel, HTML, email, EPUB, Jupyter, images, or 100+ source-code and text file types); converting between formats…

27 3d ago A 181 tokens original MIT

mineru-pdf

07

hawkongz/mineru-pdf

Skill Claude CodeCodex

High-accuracy PDF content extraction using MinerU (Shanghai AI Lab). Use this whenever the user needs to extract text, formulas, tables, or images from a complex PDF — especially academic papers, multi-column layouts, scanned documents, or any PDF where pypdf produces garbled /Cxx formula output. Trigger on: "MinerU"…

10 2mo ago A 154 tokens original MIT

mineru-pdf

08

hawkongz/mineru-pdf

Skill Claude CodeCodex

Instructions for extracting text, formulas, tables, and images from difficult PDF files with MinerU, a document-reading tool. It is intended for papers, multi-column layouts, and scanned documents where basic PDF readers may produce poor results.

10 2mo ago A 180 tokens original MIT

ai-hwp-reader

09

renovys/ai-hwp-reader

Skill Claude CodeCodex

An AI skill that tells a coding agent how to read Korean HWP and HWPX word-processing files, including files inside ZIP archives. It extracts the content and then uses it to complete the user's requested task.

0 9d ago C 0 tokens original MIT

ai-hwp-reader

10

renovys/ai-hwp-reader

Skill Claude CodeCodex

A read-only tool for opening Korean HWP and HWPX word-processing files, or ZIP archives containing them, without Hancom Office. It converts the contents to Markdown while preserving document structure.

0 9d ago A 154 tokens original MIT