Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add pantheon-org/tekhne --skill triage-papergit clone --depth 1 https://github.com/pantheon-org/tekhneWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/pantheon-org/tekhne/triage-paper)<a href="https://agentmods.dev/skills/pantheon-org/tekhne/triage-paper"><img src="https://agentmods.dev/badge/skills/pantheon-org/tekhne/triage-paper/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/pantheon-org/tekhne/triage-paper"><img src="https://agentmods.dev/badge/skills/pantheon-org/tekhne/triage-paper.svg" alt="Reviewed on agentmods" width="80" height="20"></a>- NVIDIA SkillSpector warn
SkillSpector: 1 finding, up to medium
These are SkillSpector’s own severities. On a checked sample its high-severity flags on skills were ~96% false positives — a documented command, a public API, a “never do X” rule — so we show them as a caution to read, not a verdict. Why →
- medium Excessive Agency · line 170 Skill enables autonomous high-impact decisions without human-in-the-loop verification. Critical operations (destructive commands, financial transactions, data deletion) should require explicit user confirmation.Fix: Add human-in-the-loop confirmation for destructive, irreversible, or high-impact operations. Never auto-execute commands that modify files, send data, or alter system state.
What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00090 | $0.01874 |
| Opus 5 | $0.00045 | $0.00937 |
| Sonnet 5 | $0.00018 | $0.00375 |
| Haiku 4.5 | $0.00009 | $0.00187 |
Grade A, and why
triage-paper scanned grade A with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 9d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Makes network callslowCapability
Not a fault in itself. Listed so you know the mod talks to something, and to what.
curl -s "https://arxiv.org/abs/<id>" How it starts
The opening of the file, as written. The whole thing — 182 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Triage Paper
Add a new academic paper to the research repo as a structured reference summary.
When to Use
- User provides an arxiv ID (e.g.
2310.08560), arxiv URL, or paper PDF path - User says "triage this paper", "add this paper", or "analyse this paper"
- Evaluating whether a paper belongs in the repo
When Not to Use
- The paper has already been triaged (check
REVIEWED.mdfirst) - The paper is clearly out of scope (not related to the research domain)
- User wants a full deep-dive analysis — use
triage-paperfirst, then promote toANALYSIS-*.md
Recommended MCP Servers
When available, prefer these MCPs over WebFetch for paper discovery and metadata resolution — they return structured data and avoid HTML scraping.
{
"mcpServers": {
"semantic-scholar": {
"type": "stdio",
"command": "uvx",
"args": ["semantic-scholar-fastmcp"]
},
"google-scholar": {
"type": "stdio",
"command": "uvx",
"args": ["google_scholar_mcp_server"]
}
}
}
Use semantic-scholar as the primary source (open, structured, covers most CS/ML papers). Fall back to google-scholar for papers not indexed there. Fall back to WebFetch (arxiv abstract page) only when neither MCP is configured or returns results.
Mindset
Triage is a quality gate, not a data-entry task. The goal is a scannable, honest record.
- Evidence first: quote what the paper reports; never infer or embellish claims.
- Triage-then-promote: every paper enters via
REVIEWED.md; promotion toANALYSIS-*.mdrequires a deliberate user decision — never automatic. - Scope over completeness: a well-reasoned rejection is as valuable as a full summary. If the paper is tangentially related, triage it and flag it; don't silently skip it.
Workflow
1. Resolve the source
- If given an arxiv ID or URL: use the
semantic-scholarMCP to resolve metadata (title, authors, date, abstract, DOI). If not configured, fall back toWebFetchon the arxiv abstract page. - If given a DOI: prefer
semantic-scholarorgoogle-scholarMCP over a raw HTTP fetch. - If given a PDF path, read it to extract the same fields.
- Derive a stable slug:
<firstauthor-surname>-<2-3-word-topic>(e.g.jiang-llmlingua,press-longchat).
What ships with it
15 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
- .audits/2026-04-08/analysis.md 1.2 KB
- .audits/2026-04-08/audit.json 498 B
- .audits/2026-04-08/remediation-plan.md 2.5 KB
- .audits/latest 10 B
- assets/schemas/analysis-paper.schema.json 1.3 KB
- assets/schemas/reference-paper.schema.json 1.3 KB
- assets/templates/ANALYSIS-paper.yaml 3.3 KB
- assets/templates/REFERENCE-paper.yaml 3.6 KB
- evals/instructions.json 3.2 KB
- evals/scenario-01.md 4.4 KB
- evals/scenario-02.md 4.0 KB
- evals/scenario-03.md 4.9 KB
- evals/summary.json 309 B
- scripts/validate-analysis-paper.sh 2.2 KB runs code
- scripts/validate-reference-paper.sh 2.7 KB runs code
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 9d ago First seen · 182 lines · 90 tokens per session scan A d39ab339ca7a
triage-paper is a skill published in the GitHub repository pantheon-org/tekhne (10 stars, last pushed yesterday), licensed MIT. It adds 90 tokens to every session and 1,874 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 1 finding (makes network calls). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.
Other skills, from other repositories
calculator
Evaluate mathematical expressions and unit conversions. Handles arithmetic, percentages, exponents, and common unit conversions (temperature, distance, weight). No external dependencies.
figure-rhetoric
Evaluate whether figures and plots in a manuscript effectively communicate the claims they support. Audits chart-type fit, axis design, visual hierarchy, data density, caption interpretation, perceptual accuracy, and narrative arc across 8 dimensions. Triggers on: "do my figures work", "check my plots", "are my graphs…
manuscript-provenance
Computational provenance audit verifying every number, table, and figure in a manuscript derives from code, not manual entry. Triggers on: "check provenance", "verify reproducibility", "audit my pipeline", "are my numbers from code", "provenance audit". Companion to manuscript-review (prose audit).
arxiv-package
Package a TeX/LaTeX project into a clean tarball or zip for arXiv upload: file selection, build-artifact exclusion, 00README.XXX generation, ancillary file organization, archive validation. Triggers on: "package for arXiv", "create arXiv tarball", "bundle submission", "zip for arXiv", "prepare arXiv upload", "arXiv…
paper-planning
Guides pre-writing planning for academic papers with 4 structured steps: story design (task-challenge-insight-contribution-advantage), experiment planning (comparisons + ablations), figure design (pipeline + teaser), and 4-week timeline management. Includes counterintuitive planning tactics (write a mock rejection…
iterate-ml-experiment
Owns the iteration loop on top of an ML workspace: the journal/JOURNAL.md index and the per-experiment journal/NNshortname.md design notes that must be drafted and approved by the user before experiments/NNshortname.py is created. Drives the propose → iterate → approve → implement → record loop; dispatches to…