VCVinh/Awesome-Research-Development

A rigor toolkit for anyone doing real research or R&D, plus the human-AI delegation and self-improvement research behind building agentic systems responsibly. 3 skills for Claude Code, all grounded in real 2026 papers.

1Stars on the repository
6Mods indexed here, across every type
1mo agoLast push, which is what freshness is scored on
noneNo LICENSE: all rights reserved, so bodies are not copied

VCVinh/Awesome-Research-Development

Skill Claude CodeCodex

Where the delegation boundary sits between what Claude can decide/do and what only Vinh can know or authorize — grounded in 2026 human-AI teaming/delegation research (Google DeepMind's Intelligent AI Delegation, the Interposition Problem, HAIF, the Delegated-Autonomy Boundary paper, medical-AI oversight conditions…

not rated 1 1mo ago A 113 tokens

VCVinh/Awesome-Research-Development

Skill Claude CodeCodex

How to actually make multi-agent debate/council/board deliberation improve decisions, grounded in 2026 research — not just running N perspectives and assuming more voices means better answers. Covers why vanilla multi-agent debate often underperforms simple majority vote, which decision protocol (voting vs. consensus)…

not rated 1 1mo ago A 132 tokens

VCVinh/Awesome-Research-Development

Skill Claude CodeCodex

General-purpose infrastructure for any autonomous, large-scale research pipeline — not specific to one project. Use whenever building, running, or trusting a pipeline that searches many sources (dozens to thousands), digests them, deeply comprehends them (not just restates them), and produces the widest plausible set…

not rated 1 1mo ago A 0 tokens

VCVinh/Awesome-Research-Development

Skill Claude CodeCodex

Trending, real (arXiv-grounded, 2025-2026) techniques for continual/lifelong learning, agent harness engineering, and recursive self-improvement (RSI). Use when designing an agent that must keep getting better across sessions without catastrophic forgetting, when optimizing the harness…

not rated 1 1mo ago A 158 tokens

VCVinh/Awesome-Research-Development

Skill Claude CodeCodex

When and how to put a semantic layer (Cube, dbt MetricFlow, or bespoke) between an LLM agent and a data warehouse — governed metrics/dimensions/access-rules instead of letting an agent invent raw SQL. Use when an agent needs to query financial/business data (revenue, positions, user metrics) and the same question must…

not rated 1 1mo ago A 130 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: