← ClaudeAtlas

ai-agents-research-frontierlisted

Three ranked open research programs for this repo, each with honest current-state evidence, first concrete steps, and a falsifiable milestone. Verified governance (ADR-069, proposed), cross-harness abstraction (ADR-072 proposed, ADR-068 accepted), and the self-improving loop (issue #1345). Use when you say `research frontier`, `open problems`, `what should we research next`. Do NOT use for how to run an experiment here (use `ai-agents-research-methodology`).
rjmurillo/ai-agents · ★ 41 · AI & Automation · score 77
Install: claude install-skill rjmurillo/ai-agents
# AI Agents Research Frontier <!-- vendor-portability: contributor-facing knowledge pack for the rjmurillo/ai-agents repo itself; intentionally references upstream paths (.agents/, .claude/, scripts/, build/) because its audience is repo contributors, not plugin consumers (issue #2050) --> Open problems where this repository can advance the state of the art, ranked in owner-confirmed priority order. This skill tells you WHAT is worth working on and what "done" would look like. It does not teach experiment discipline (that is `ai-agents-research-methodology`) and it is not the portability battle plan (that is `ai-agents-portability-campaign`). Honesty contract for this document: every asset claim below was re-verified against the working tree on 2026-07-30. Anything labeled PROPOSED is not policy. Anything labeled UNVERIFIED could not be confirmed and must be checked before you build on it. ## Triggers - `research frontier` - `open problems` - `what should we research next` - `frontier programs` ## The Three Programs at a Glance | Rank | Program | Anchor artifacts | Status (as of 2026-07-30) | Falsifiable milestone (short form) | |------|---------|------------------|---------------------------|-------------------------------------| | 1 | Verified governance | ADR-069, `scripts/eval/eval-rule-activation.py` | ADR-069 is PROPOSED; eval tool exists, 15 rule scenario fixtures exist | Controlled eval shows gated-corpus sessions beat ungated on N scenarios with defensible stat