pdfium skills

26 tagged pdfium, measured the same way as everything else here.

Browse within: document-intelligence 26metadata-extraction 26pdf-extraction 26text-extraction 26

xberg-io/xberg

Skill Claude CodeCodex

Design, implement, or diagnose Xberg plugin traits, typed registries, priority collisions, lifecycle, native extractors, and Alef-generated Python plugin bridges. Load for plugin-system work, not ordinary extractor parsing.

9.2k +12 yesterday A 49 tokens original MIT

extracting-keywords

02

xberg-io/xberg

Skill Claude CodeCodex

Part of xberg

Use when extracting keywords (YAKE/RAKE) from documents — and, secondarily, when detecting document language or generating embeddings for RAG and search. Covers the keyword config (and its feature gating), --detect-language, and the standalone embed command with real flags.

9.2k +12 yesterday A 63 tokens original MIT

xberg

03

xberg-io/xberg

Skill Claude CodeCodex

Part of xberg

Extract text, tables, metadata, and images from 107 document formats (PDF, Office, images, HTML, email, archives, academic) using Xberg. Use when writing code that calls Xberg APIs in Python, Node.js/TypeScript, Rust, or CLI. Covers installation, extraction (sync/async), configuration (OCR, chunking, output format)…

9.2k +12 changed today A 87 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: