paper-analyzer

paper-analyzer is a skill for Claude Code from proyecto26/sherlock-ai-plugin. It costs 55 tokens per session (659 once invoked), scanned A, original, MIT.

A tool for turning academic papers into technical articles by parsing PDFs and extracting text, images, tables, and mathematical formulas.

In plain words
What is it for?
Use it to create Markdown or HTML articles in different writing styles, explain formulas, and optionally connect the discussion to open-source code on GitHub.
Why use it?
It reduces the manual work of reading complex papers and preparing their contents in reusable document formats.

Skill for Claude Code

Written for Claude Code: shipped in a Claude Code plugin.

Part of the sherlock-ai-plugin plugin — 6 skills shipped together

Good fit Use it to create Markdown or HTML articles in different writing styles, explain formulas, and optionally connect the discussion to open-source code on GitHub.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/proyecto26/sherlock-ai-plugin/paper-analyzer
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add proyecto26/sherlock-ai-plugin --skill paper-analyzer
Clone the repo
git clone --depth 1 https://github.com/proyecto26/sherlock-ai-plugin

Made for: Claude Code.

Or install sherlock-ai-plugin, the plugin that ships this one along with the rest of its 6 skills.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for paper-analyzer

README.md
[![agentmods](https://agentmods.dev/badge/skills/proyecto26/sherlock-ai-plugin/paper-analyzer.svg)](https://agentmods.dev/skills/proyecto26/sherlock-ai-plugin/paper-analyzer)
Your own site
<a href="https://agentmods.dev/skills/proyecto26/sherlock-ai-plugin/paper-analyzer"><img src="https://agentmods.dev/badge/skills/proyecto26/sherlock-ai-plugin/paper-analyzer.svg" alt="Measured on agentmods" height="20"></a>
Per session 55 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 659 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00055 $0.00659
Opus 5 $0.00028 $0.00329
Sonnet 5 $0.00011 $0.00132
Haiku 4.5 $0.00006 $0.00066

Measured 7d ago against content hash 972483dee542, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-07, from the pricing page.

Security

Grade A, and why

paper-analyzer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 7d ago.

The scan reads SKILL.md. This mod also ships 4 executable files (scripts/convert_pdf.py, scripts/extract_paper_info.py, scripts/generate_html.py, …), listed below but not scanned — reading those needs a real analyzer, not pattern matching.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/paper-analyzer/SKILL.md · 95 lines

How it starts

The opening of the file, as written. The whole thing — 95 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Academic Paper Analyzer – In-Depth Analysis of Academic Papers

Core Capabilities

  • MinerU Cloud API for high-precision PDF parsing
  • Automatic extraction of images, tables, and LaTeX formulas
  • Multiple writing styles: storytelling / academic / concise
  • Optional formula explanations: insert formula images with detailed symbol explanations
  • Optional code analysis: combine explanations with GitHub open-source code
  • Output Markdown + HTML (base64-embedded images)

Prerequisites

MinerU API Token

  1. Visit https://mineru.net and register an account
  2. Obtain an API Token
  3. Set an environment variable (recommended):
    export MINERU_TOKEN="your_token_here"
    

Dependency Installation

pip install requests markdown

Workflow

Step 1: PDF Parsing (Using MinerU API)

python scripts/mineru_api.py <pdf_path> <output_dir>

Or pass the token directly:

python scripts/mineru_api.py paper.pdf ./output YOUR_TOKEN

Output:

  • output_dir/*.md – Markdown files (including formulas and tables)
  • output_dir/images/ – High-quality extracted images

Step 2: Extract Paper Metadata

python scripts/extract_paper_info.py <output_dir>/*.md paper_info.json

Step 3: Style Selection (Ask the User)

Before generating the article, you must ask the user to choose the following options:

1. Writing Style (Required)
Style Characteristics Use Cases
storytelling Starts from intuition, uses metaphors and examples, narrative-driven Blogs, tech columns, popular science
academic Professional terminology, rigorous expression, preserves original concepts Academic reports, surveys, research group sharing
concise Straight to the point, tables and lists, high information density Quick reads, paper overviews, technical research
2. Formula Option (Optional)
Option Description
with-formulas Insert formula images and explain symbol meanings in detail
no-formulas (default) Pure text description, no formula images

Read the full file on GitHub · 95 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 7d ago First seen · 95 lines · 55 tokens per session scan A 972483dee542

Subscribe to this mod's changes

paper-analyzer is a skill published in the GitHub repository proyecto26/sherlock-ai-plugin (35 stars, last pushed 6d ago), licensed MIT. It adds 55 tokens to every session and 659 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

paper-compile

A build workflow that turns LaTeX source files into a PDF and checks whether the paper compiles correctly. LaTeX is a text-based system commonly used for academic papers.

wanshuiyin/Auto-claude-code-research-in-sleep · 53 tokens

ccs-submission

Use when auditing an ACM CCS submission for HotCRP readiness, dual-cycle abstract registration and full-paper deadlines, the 12-page ACM sigconf body, anonymization, the ethics considerations appendix, artifact-availability posture, dual-submission policy, per-cycle submission caps, desk-reject triggers, and…

brycewang-stanford/Awesome-Journal-Skills · 69 tokens

acmmm-camera-ready

Use when preparing the ACM MM (ACM Multimedia) camera-ready version of record — de-anonymizing safely, completing the ACM rights form and CCS concepts, meeting ACM sigconf requirements, releasing code/data/media artifacts and any earned reproducibility badge, registering, and planning the oral/poster presentation in…

brycewang-stanford/Awesome-Journal-Skills · 67 tokens

fin-paper-convert

Compile LaTeX to PDF and convert to target journal format.

csmar432/finai-research · 12 tokens

beamer-deck

Create an academic presentation as a LaTeX Beamer source and reviewed PDF with an original theme. Use when the requested deliverable is a conference, seminar, or lecture deck in Beamer. Not for PowerPoint or RevealJS; use $pptx or $quarto-deck.

flonat/flonat-research · 63 tokens

bib-parse

Extract citations from a PDF and generate a validated .bib file. Use when the user asks to extract citations from a PDF and generate a validated .bib file. Reads the PDF, identifies referenced works, constructs BibTeX entries, and verifies metadata.

flonat/flonat-research · 54 tokens