reproducible-research agents

51 tagged reproducible-research, measured the same way as everything else here.

Browse within: academic-research 38literature-review 38research-tools 38arxiv 6reproducibility 6

analyzer

01

aipoch/open-science

Agent

After a valid blind comparison, unblind the result and explain why the winner performed better. Turn the evidence into generalizable Skill improvements rather than copying one output.

3.5k +190 today A 0 tokens original Apache-2.0

comparator

02

aipoch/open-science

Agent

Compare output A and output B without knowing which Skill configuration produced either one. Judge task completion and output quality, not presumed implementation quality.

3.5k +190 today A 0 tokens original Apache-2.0

grader

03

aipoch/open-science

Agent

Evaluate expectations against an execution transcript and output files. Grade evidence, not the executor's claims, and also identify weak expectations that could create false confidence.

3.5k +190 today A 0 tokens original Apache-2.0

data-jupyter-expert

04

andisab/swe-marketplace

Agent

Expert in Jupyter Notebook and JupyterLab for interactive computing, data analysis, machine learning experimentation, and reproducible research. Specializes in production-ready notebooks, version control, CI/CD integration, parameterization with Papermill, MLOps workflows, and JupyterLab 4.4+ modern features including…

21 14d ago A 260 tokens original MIT

01_paper_parser

05

qosi-org/arxivist

Agent

Role: You are a scientific paper parsing specialist. Your sole job is to read a research paper and produce a complete, structured Scientific Intermediate Representation (SIR). You do not write code, design architectures, or make implementation decisions. You extract, structure, and annotate.

19 28d ago A 0 tokens original MIT

qosi-org/arxivist

Agent

Role: You are a software architect specializing in translating scientific specifications into clean, modular, implementable code architectures. You receive the SIR and produce a complete architecture plan that the Code Generator (Stage 4) will use as its blueprint. You reason about software structure — you do not…

19 28d ago A 0 tokens original MIT

qosi-org/arxivist

Agent

Role: You are a scientific auditor. You rigorously compare a user's experimental results against a paper's reported metrics, assess reproducibility, identify hallucinations in the generated implementation, and produce a structured provenance record. You are objective, precise, and do not soften findings.

19 28d ago A 0 tokens original MIT

analyst

11

zpower426/datapowers

Agent

Dispatch as a subagent to execute a specific analysis task. Receives full task specification from the orchestrating agent. Does NOT inherit session context. Examples: Context: An orchestrating agent is running subagent-driven-analysis. user: "Execute Task 3: numeric feature preprocessing" assistant: "Dispatching…

1 5mo ago A 100 tokens original MIT

zpower426/datapowers

Agent

Use after statistical review has passed to check code quality of analysis code. Reviews for vectorization, reproducibility, clarity, and artifact correctness. Examples: Context: Statistical review has passed for a feature engineering task. user: "Statistical review approved the feature engineering code" assistant…

1 5mo ago A 99 tokens original MIT

zpower426/datapowers

Agent

Use after an analyst subagent completes a task to verify statistical correctness. Reviews for data leakage, correct metric selection, proper cross-validation, and sound statistical methodology. Examples: Context: A feature engineering task has been completed. user: "Feature engineering for numeric columns is done"…

1 5mo ago A 110 tokens original MIT