Deep evidence-first research with broad discovery, verification, and traceable citations. Prefer invoking via research-workflow. TRIGGER when (MANDATORY — you MUST invoke this skill, no exceptions): user message contains ANY of these keywords or synonyms — 调研/研究/对比/综述/文献/证据/机制/根因/为什么/可行性/路线图/分析/探索, or…
Execute AI/ML experiments locally or remotely with environment, runtime, and logging controls. Prefer invoking via research-workflow. TRIGGER when: user asks to run/launch/start/resume/monitor a training job, evaluation, or benchmark, or a plan is ready for execution, or experiment needs rerun/recovery. DO NOT TRIGGER…
Escalate critical decisions to a human with options, tradeoffs, and recommendation. TRIGGER when: major safety risk, high irreversible impact, large GPU spend, destructive data ops, shared-memory publication, or hard blockers requiring human approval. DO NOT TRIGGER when: routine confirmations handled by run-governor…
Manage long-term AI R&D memory: retrieval, writeback, promotion, and shared export. TRIGGER when: run bootstrap, each new user turn, each execution batch, significant failure, replan, high-resource action, long-action resume, final report handoff, or compaction markers detected (Compact/压缩/Summary). DO NOT TRIGGER…
Write CS/AI papers with progressive disclosure. Prefer invoking via research-workflow. TRIGGER when: user explicitly asks to draft/write/revise a paper, paper section (abstract, intro, related work, method, experiments, conclusion), rebuttal, or LaTeX content. DO NOT TRIGGER when: user mentions papers but needs…
Initialize and maintain per-project runtime context (env, secrets, snapshots). Prefer invoking via research-workflow. TRIGGER when: new run needs env setup, preflight before experiment/eval, runtime fields missing (paths, API keys, GPU config, proxy), run snapshot needed, or shared-memory needs project config. DO NOT…
Build execution-ready research plans, experiment proposals, and ablation designs. Prefer invoking via research-workflow. TRIGGER when: user asks for research proposal, experiment plan, ablation design, evaluation roadmap, study design, or after deep-research scoping needs conversion to actionable plan. DO NOT TRIGGER…
PRIMARY ORCHESTRATOR — trigger this skill FIRST for any non-trivial AI R&D task. Coordinates run-governor, memory-manager, deep-research, research-plan, project-context, experiment-execution, human-checkpoint, and paper-writing. TRIGGER FIRST when: any non-trivial research task begins (analysis, debugging…
Govern run-level execution policy: mode selection, durable run tracking, long-action watch/resume policy, stage reporting, and safety allowances. TRIGGER when: starting a non-trivial research task (set mode + runid), switching local/remote target, creating a new run, or mode-aware policy decisions needed. DO NOT…