Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add karaage0703/ai-assistant-workspace --skill xs-arxivgit clone --depth 1 https://github.com/karaage0703/ai-assistant-workspaceWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/karaage0703/ai-assistant-workspace/xs-arxiv)<a href="https://agentmods.dev/skills/karaage0703/ai-assistant-workspace/xs-arxiv"><img src="https://agentmods.dev/badge/skills/karaage0703/ai-assistant-workspace/xs-arxiv/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/karaage0703/ai-assistant-workspace/xs-arxiv"><img src="https://agentmods.dev/badge/skills/karaage0703/ai-assistant-workspace/xs-arxiv.svg" alt="Reviewed on agentmods" width="80" height="20"></a>- NVIDIA SkillSpector pass
What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00079 | $0.02809 |
| Opus 5 | $0.00039 | $0.01404 |
| Sonnet 5 | $0.00016 | $0.00562 |
| Haiku 4.5 | $0.00008 | $0.00281 |
Grade A, and why
xs-arxiv scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 244 lines — stays where its author put it; the contents beside it link to each section on GitHub.
arXiv論文調査スキル
重要:デフォルトはabstractのみ
- 全文読み込みはデフォルトでやらない。 abstractベースで要約・スコアリングする
- 全文読み込みをするのは「この論文詳しく」「全文読んで」と明示的に言われた時だけ
- 検索コマンドは1回で済ませる(複数クエリを投げない)
- 中間報告を出す — 検索開始時に「arXiv検索中…」と一言送ってからコマンド実行
興味度スコアの判断基準
ユーザーの興味分野は AGENTS.md の「ユーザーについて」セクションを参照する。未設定の場合は以下のデフォルトカテゴリで判断:
- 最高関心(★★★): LLM/プロンプトエンジニアリング、AIエージェント、RAG、Computer Use/ブラウザ操作
- 高関心(★★☆): ロボティクス/Embodied AI、時系列予測、エッジAI/ローカル推論、TTS/音声合成
- 関心あり(★☆☆): 勾配ブースティング、コンピュータビジョン、マルチモーダル、3D生成
ユーザー側で AGENTS.md に固有の興味分野(例: 量子情報、創薬AI、強化学習)を書いていれば、そちらを優先する。
実行フロー
Step 1: モード判定
- 検索モード — 特定トピックの論文を探す
- トレンドモード — 最新の注目論文を発見する(スケジュール実行はこれ)
- 分析モード — 特定論文を深く読む
Step 2: 論文検索・トレンド取得
3 経路ある。用途に応じて使い分ける (export.arxiv.org API は混雑時間帯で 429/503 を返しがちで非推奨、fallback 扱い)。
A. トレンドモード (毎朝の自動実行はこれ)
arxiv 公式 RSS から並列取得。レート制限ほぼなし、1〜2 秒で完了。
cd [SKILL_DIR]/scripts && uv run python arxiv_tool.py trending -c cs.AI cs.LG cs.CL cs.CV -d 7 -n 20
返り値: {"total_results": N, "papers": [...], "categories_queried": [...], "categories_failed": [...], "source": "rss"}
B. 検索モード (任意クエリでの探索)
Semantic Scholar API がデフォルト。abstract / 著者 / 引用数 / OA PDF URL までリッチに取れる。
cd [SKILL_DIR]/scripts && uv run python arxiv_tool.py search "transformer attention" -n 10 [--year-from 2025]
返り値: {"total_results": N, "papers": [{id, title, authors, abstract, published, citation_count, ...}], "source": "s2"}
C. fallback: export.arxiv.org API (legacy)
旧経路。S2 が応答しない・S2 にない最新論文を狙うときだけ。指数バックオフリトライ済 (15s→45s→90s)。
cd [SKILL_DIR]/scripts && uv run python arxiv_tool.py search "クエリ" -n 10 --source legacy --date-from 2026-05-01 -s date
検索クエリの最適化(B/C 共通)
- 引用句でフレーズ検索:
"multi-agent systems" - OR演算子で関連技術をカバー:
"AI agents" OR "intelligent agents" - フィールド指定検索:
ti:"exact title",au:"author name",abs:"keyword" - 除外検索:
"machine learning" ANDNOT "survey"
主要カテゴリ: cs.AI / cs.LG / cs.CL / cs.CV / cs.MA / cs.RO
エラー時の挙動
S2 や legacy が全リトライ尽きると JSON が以下の形:
What ships with it
7 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 12d ago First seen · 244 lines · 79 tokens per session scan A 4c9527fc25e3
xs-arxiv is a skill published in the GitHub repository karaage0703/ai-assistant-workspace (137 stars, last pushed 25d ago), licensed MIT. It adds 79 tokens to every session and 2,809 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
instrument-data-to-allotrope
Convert laboratory instrument output files (PDF, CSV, Excel, TXT) to Allotrope Simple Model (ASM) JSON format or flattened 2D CSV. Use this skill when scientists need to standardize instrument data for LIMS systems, data lakes, or downstream analysis. Supports auto-detection of instrument types. Outputs include full…
matlab
Build, review, migrate, and safely plan MATLAB or GNU Octave numerical workflows, including arrays, tabular/time data, tests, projects, graphics, MAT files, and explicit Python interoperability.
exploratory-data-analysis
Perform bounded, local exploratory analysis of explicitly supported scientific files. Use for redacted CSV/TSV/JSON profiles; optional NumPy, HDF5, FASTA/FASTQ, and basic image metadata inspection; missingness/leakage audits; outlier and transformation sensitivity; and rigorous EDA report scaffolds. Other domain…
phylogenetics
Build and analyze phylogenetic trees using MAFFT (multiple alignment), IQ-TREE 2 (maximum likelihood), and FastTree (fast NJ/ML). Visualize with ETE3 or FigTree. For evolutionary analysis, microbial genomics, viral phylodynamics, protein family analysis, and molecular clock studies.
research-engineer
An uncompromising Academic Research Engineer. Operates with absolute scientific rigor, objective criticism, and zero flair. Focuses on theoretical correctness, formal verification, and optimal implementation across any required technology.
mapping-to-snomed
Maps clinical concept spans extracted by OpenMed to SNOMED CT concepts through a USER-SUPPLIED terminology server (the user's own Ontoserver, Snowstorm, or UMLS/UTS), never a bundled vocabulary. Use when the user wants to code findings, disorders, procedures, body structures, or substances to SNOMED CT, run an ECL…