xs-arxiv

xs-arxiv is a skill for Claude Code, Codex from karaage0703/ai-assistant-workspace. It costs 79 tokens per session (2,809 once invoked), scanned A, original, MIT.

A research workflow for finding and summarizing papers on arXiv, an online repository for research preprints, using their abstracts by default.

In plain words
What is it for?
Use it to search topics, collect recent papers in selected fields, find trending work, or perform a deeper analysis when the user explicitly asks for the full text.
Why use it?
It provides a consistent way to search papers, identify trends, and judge relevance without reading every full paper unless requested.

Skill for Claude CodeCodex

Written for no agent in particular: nothing here depends on one. Also seen: mentions AGENTS.md.

Good fit Use it to search topics, collect recent papers in selected fields, find trending work, or perform a deeper analysis when the user explicitly asks for the full text.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/karaage0703/ai-assistant-workspace/xs-arxiv
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add karaage0703/ai-assistant-workspace --skill xs-arxiv
Clone the repo
git clone --depth 1 https://github.com/karaage0703/ai-assistant-workspace

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for xs-arxiv

README.md
[![agentmods](https://agentmods.dev/badge/skills/karaage0703/ai-assistant-workspace/xs-arxiv/github.svg)](https://agentmods.dev/skills/karaage0703/ai-assistant-workspace/xs-arxiv)
Your own site
<a href="https://agentmods.dev/skills/karaage0703/ai-assistant-workspace/xs-arxiv"><img src="https://agentmods.dev/badge/skills/karaage0703/ai-assistant-workspace/xs-arxiv/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for xs-arxiv

Your own site · 80×15
<a href="https://agentmods.dev/skills/karaage0703/ai-assistant-workspace/xs-arxiv"><img src="https://agentmods.dev/badge/skills/karaage0703/ai-assistant-workspace/xs-arxiv.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 79 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,809 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe. Third-party audits
  • NVIDIA SkillSpector pass 7 Sept 2026
How audits are shown
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00079 $0.02809
Opus 5 $0.00039 $0.01404
Sonnet 5 $0.00016 $0.00562
Haiku 4.5 $0.00008 $0.00281

Measured 12d ago against content hash 4c9527fc25e3, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-12, from the pricing page.

Security

Grade A, and why

xs-arxiv scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.

The scan reads SKILL.md. This mod also ships 4 executable files (scripts/arxiv_fetcher.py, scripts/arxiv_tool.py, scripts/rss_provider.py, …), listed below but not scanned — reading those needs a real analyzer, not pattern matching.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/xs-arxiv/SKILL.md · 244 lines

How it starts

The opening of the file, as written. The whole thing — 244 lines — stays where its author put it; the contents beside it link to each section on GitHub.

arXiv論文調査スキル

重要:デフォルトはabstractのみ

  • 全文読み込みはデフォルトでやらない。 abstractベースで要約・スコアリングする
  • 全文読み込みをするのは「この論文詳しく」「全文読んで」と明示的に言われた時だけ
  • 検索コマンドは1回で済ませる(複数クエリを投げない)
  • 中間報告を出す — 検索開始時に「arXiv検索中…」と一言送ってからコマンド実行

興味度スコアの判断基準

ユーザーの興味分野は AGENTS.md の「ユーザーについて」セクションを参照する。未設定の場合は以下のデフォルトカテゴリで判断:

  • 最高関心(★★★): LLM/プロンプトエンジニアリング、AIエージェント、RAG、Computer Use/ブラウザ操作
  • 高関心(★★☆): ロボティクス/Embodied AI、時系列予測、エッジAI/ローカル推論、TTS/音声合成
  • 関心あり(★☆☆): 勾配ブースティング、コンピュータビジョン、マルチモーダル、3D生成

ユーザー側で AGENTS.md に固有の興味分野(例: 量子情報、創薬AI、強化学習)を書いていれば、そちらを優先する。

実行フロー

Step 1: モード判定

  • 検索モード — 特定トピックの論文を探す
  • トレンドモード — 最新の注目論文を発見する(スケジュール実行はこれ)
  • 分析モード — 特定論文を深く読む

Step 2: 論文検索・トレンド取得

3 経路ある。用途に応じて使い分ける (export.arxiv.org API は混雑時間帯で 429/503 を返しがちで非推奨、fallback 扱い)。

A. トレンドモード (毎朝の自動実行はこれ)

arxiv 公式 RSS から並列取得。レート制限ほぼなし、1〜2 秒で完了。

cd [SKILL_DIR]/scripts && uv run python arxiv_tool.py trending -c cs.AI cs.LG cs.CL cs.CV -d 7 -n 20

返り値: {"total_results": N, "papers": [...], "categories_queried": [...], "categories_failed": [...], "source": "rss"}

B. 検索モード (任意クエリでの探索)

Semantic Scholar API がデフォルト。abstract / 著者 / 引用数 / OA PDF URL までリッチに取れる。

cd [SKILL_DIR]/scripts && uv run python arxiv_tool.py search "transformer attention" -n 10 [--year-from 2025]

返り値: {"total_results": N, "papers": [{id, title, authors, abstract, published, citation_count, ...}], "source": "s2"}

C. fallback: export.arxiv.org API (legacy)

旧経路。S2 が応答しない・S2 にない最新論文を狙うときだけ。指数バックオフリトライ済 (15s→45s→90s)。

cd [SKILL_DIR]/scripts && uv run python arxiv_tool.py search "クエリ" -n 10 --source legacy --date-from 2026-05-01 -s date
検索クエリの最適化(B/C 共通)
  • 引用句でフレーズ検索: "multi-agent systems"
  • OR演算子で関連技術をカバー: "AI agents" OR "intelligent agents"
  • フィールド指定検索: ti:"exact title", au:"author name", abs:"keyword"
  • 除外検索: "machine learning" ANDNOT "survey"

主要カテゴリ: cs.AI / cs.LG / cs.CL / cs.CV / cs.MA / cs.RO

エラー時の挙動

S2 や legacy が全リトライ尽きると JSON が以下の形:

Read the full file on GitHub · 244 lines

Files

What ships with it

7 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 12d ago First seen · 244 lines · 79 tokens per session scan A 4c9527fc25e3

Subscribe to this mod's changes

xs-arxiv is a skill published in the GitHub repository karaage0703/ai-assistant-workspace (137 stars, last pushed 25d ago), licensed MIT. It adds 79 tokens to every session and 2,809 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

instrument-data-to-allotrope

Convert laboratory instrument output files (PDF, CSV, Excel, TXT) to Allotrope Simple Model (ASM) JSON format or flattened 2D CSV. Use this skill when scientists need to standardize instrument data for LIMS systems, data lakes, or downstream analysis. Supports auto-detection of instrument types. Outputs include full…

anthropics/knowledge-work-plugins · 123 tokens

matlab

Build, review, migrate, and safely plan MATLAB or GNU Octave numerical workflows, including arrays, tabular/time data, tests, projects, graphics, MAT files, and explicit Python interoperability.

K-Dense-AI/scientific-agent-skills · 42 tokens

exploratory-data-analysis

Perform bounded, local exploratory analysis of explicitly supported scientific files. Use for redacted CSV/TSV/JSON profiles; optional NumPy, HDF5, FASTA/FASTQ, and basic image metadata inspection; missingness/leakage audits; outlier and transformation sensitivity; and rigorous EDA report scaffolds. Other domain…

K-Dense-AI/scientific-agent-skills · 83 tokens

phylogenetics

Build and analyze phylogenetic trees using MAFFT (multiple alignment), IQ-TREE 2 (maximum likelihood), and FastTree (fast NJ/ML). Visualize with ETE3 or FigTree. For evolutionary analysis, microbial genomics, viral phylodynamics, protein family analysis, and molecular clock studies.

K-Dense-AI/scientific-agent-skills · 68 tokens

research-engineer

An uncompromising Academic Research Engineer. Operates with absolute scientific rigor, objective criticism, and zero flair. Focuses on theoretical correctness, formal verification, and optimal implementation across any required technology.

davila7/claude-code-templates · 43 tokens

mapping-to-snomed

Maps clinical concept spans extracted by OpenMed to SNOMED CT concepts through a USER-SUPPLIED terminology server (the user's own Ontoserver, Snowstorm, or UMLS/UTS), never a bundled vocabulary. Use when the user wants to code findings, disorders, procedures, body structures, or substances to SNOMED CT, run an ECL…

maziyarpanahi/openmed · 205 tokens