ai for science agents

29 tagged ai for science, measured the same way as everything else here.

Browse within: llm-agents 12amd 11drug-discovery 11earth-science 11fine-tuning 11healthcare 11data-analysis 10instruments 10magnetism 10autonomous-research 5

analyzer

01

aipoch/open-science

Agent

After a valid blind comparison, unblind the result and explain why the winner performed better. Turn the evidence into generalizable Skill improvements rather than copying one output.

3.3k yesterday A 0 tokens original Apache-2.0

comparator

02

aipoch/open-science

Agent

Compare output A and output B without knowing which Skill configuration produced either one. Judge task completion and output quality, not presumed implementation quality.

3.3k yesterday A 0 tokens original Apache-2.0

grader

03

aipoch/open-science

Agent

Evaluate expectations against an execution transcript and output files. Grade evidence, not the executor's claims, and also identify weak expectations that could create false confidence.

3.3k yesterday A 0 tokens original Apache-2.0

comms-writer

09

TaewoooPark/MagLab

Agent

Delegate when drafting research communications, summaries, or reports for a non-specialist audience. Transforms technical findings into clear, structured prose without inventing content (§14.7).

11 1mo ago A 40 tokens original MIT

paper-reviewer

10

TaewoooPark/MagLab

Agent

Reads 3–7 key papers in depth and extracts claims, evidence, and methodology. Targets T1·T2 papers verified by citation-auditor (§14.7).

11 1mo ago A 40 tokens original MIT

search-scout

11

TaewoooPark/MagLab

Agent

Broadly collects candidate papers using MCP connectors and skills. Generates 3–6 query families and assigns tier classifications (§14.7).

11 1mo ago A 31 tokens original MIT

launcher

12

AMDResearch/ai4science-studio

Agent

Submits a 2-node ORBIT-2 training (AMD Instinct MI355X) with PyTorch profiling and Omnistat user-mode telemetry, waits for completion, and writes manifest.json for downstream subagents.

4 1mo ago A 0 tokens original MIT

omnistat_analyst

13

AMDResearch/ai4science-studio

Agent

Drive omnistat-inspect (PR #271) through the analyze-job phases on the user-mode VictoriaMetrics database, then map findings to the bottleneck taxonomy.

4 1mo ago A 0 tokens original MIT

omnistat_verifier

14

AMDResearch/ai4science-studio

Agent

Independently re-derive the top 2-3 claims from omnistat/claims.json by issuing raw PromQL via curl against the same VictoriaMetrics endpoint, at the finest sampling step. Optionally probe one cheap remedy on a 1-node interactive srun.

4 1mo ago A 0 tokens original MIT