liteparse

liteparse is a skill for Claude Code from yanjumlinnb-boop/scientific-agent-skills. It costs 98 tokens per session (2,398 once invoked), scanned A, a copy of liteparse, MIT.

A local tool for extracting text and layout information from PDFs, Word documents, office files, and images. It can preserve where text appears on a page, run OCR on scans, and create page screenshots.

In plain words
What is it for?
Use it for document search, scanned-document OCR, layout-aware retrieval, page images, and batch processing of research papers or other document folders.
Why use it?
It lets applications process documents locally without sending them to a cloud service or relying on a language model. Preserved positions help when the location of text, tables, or figures matters.

Skill for Claude Code

Written for Claude Code: allowed-tools in frontmatter.

Good fit Use it for document search, scanned-document OCR, layout-aware retrieval, page images, and batch processing of research papers or other document folders.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/yanjumlinnb-boop/scientific-agent-skills/liteparse
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add yanjumlinnb-boop/scientific-agent-skills --skill liteparse
Clone the repo
git clone --depth 1 https://github.com/yanjumlinnb-boop/scientific-agent-skills

Made for: Claude Code.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for liteparse

README.md
[![agentmods](https://agentmods.dev/badge/skills/yanjumlinnb-boop/scientific-agent-skills/liteparse/github.svg)](https://agentmods.dev/skills/yanjumlinnb-boop/scientific-agent-skills/liteparse)
Your own site
<a href="https://agentmods.dev/skills/yanjumlinnb-boop/scientific-agent-skills/liteparse"><img src="https://agentmods.dev/badge/skills/yanjumlinnb-boop/scientific-agent-skills/liteparse/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for liteparse

Your own site · 80×15
<a href="https://agentmods.dev/skills/yanjumlinnb-boop/scientific-agent-skills/liteparse"><img src="https://agentmods.dev/badge/skills/yanjumlinnb-boop/scientific-agent-skills/liteparse.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 98 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,398 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 1 finding. A grade says what 26 rules found in the file — not that it is safe.
Origin 91% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00098 $0.02398
Opus 5 $0.00049 $0.01199
Sonnet 5 $0.00020 $0.00480
Haiku 4.5 $0.00010 $0.00240

Measured 12d ago against content hash 8990aa277558, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-12, from the pricing page.

Security

Grade A, and why

liteparse scanned grade A with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.

The scan reads SKILL.md. This mod also ships 1 executable file (scripts/batch_parse_dir.py), listed below but not scanned — reading those needs a real analyzer, not pattern matching.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Makes network callslowCapability

Not a fault in itself. Listed so you know the mod talks to something, and to what.

curl -sL https://example.com/report.pdf | lit parse -
Origin

This is a copy

91% identical to liteparse — 23 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

skills/liteparse/SKILL.md · 294 lines

How it starts

The opening of the file, as written. The whole thing — 294 lines — stays where its author put it; the contents beside it link to each section on GitHub.

LiteParse — Local Document Parsing

Overview

LiteParse is a fast, open-source document parser (Rust core, Python/Node bindings) focused on local, layout-aware text extraction with bounding boxes. It does not produce Markdown and does not call cloud LLMs. Outputs are plain text (layout-preserved) or structured JSON with per-page text_items (position, font metadata, optional confidence).

Version note: Examples target liteparse 2.0.0 (PyPI, May 2026). The upstream V1 branch is legacy; this skill documents V2 / main only.

For parser selection vs MarkItDown, the pdf skill, or LlamaParse, see references/choosing_a_parser.md.

When to Use This Skill

Use LiteParse when you need:

  • Fast local parsing of PDFs or converted Office/image files without cloud dependencies
  • Spatial text with bounding boxes for layout-aware RAG, citation grounding, or figure/table region logic
  • OCR on scanned PDFs or images (bundled Tesseract, or a user-run HTTP OCR server)
  • Page screenshots (PNG) for multimodal agents that must see charts, figures, or handwriting
  • Batch ingestion of literature folders, supplementary PDFs, or protocol libraries
  • Page subsets or password-protected PDFs

When Not to Use

Task Use instead
Markdown for LLM ingestion (EPUB, audio, YouTube, HTML) markitdown skill
Merge/split PDFs, forms, watermarks, rotation pdf skill
Dense tables, handwriting, production cloud pipelines LlamaParse (cloud; sign up separately)

Installation

uv pip install "liteparse==2.0.0"

This installs the Python bindings and the lit CLI. Verify:

lit --help
python -c "import liteparse; print(liteparse.__version__)"

Optional system tools (for non-PDF inputs):

  • LibreOffice — Word, Excel, PowerPoint, OpenDocument, CSV/TSV
  • ImageMagick — PNG, JPEG, TIFF, WebP, SVG, etc.

Install commands are in references/ocr_and_formats.md.

Read the full file on GitHub · 294 lines

Files

What ships with it

6 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 12d ago First seen · 294 lines · 98 tokens per session scan A 8990aa277558

Subscribe to this mod's changes

liteparse is a skill published in the GitHub repository yanjumlinnb-boop/scientific-agent-skills (2 stars, last pushed 2mo ago), licensed MIT. It adds 98 tokens to every session and 2,398 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 1 finding (makes network calls). It is 91% identical to liteparse, differing in 23 lines, and is treated as a copy.

Related

Other skills, from other repositories