ai-agents-research-frontier

Solid

Three ranked open research programs for this repo, each with honest current-state evidence, first concrete steps, and a falsifiable milestone. Verified governance (ADR-069, proposed), cross-harness abstraction (ADR-072 and ADR-068, proposed), and the self-improving loop (issue #1345). Use when you say `research frontier`, `open problems`, `what should we research next`. Do NOT use for how to run an experiment here (use `ai-agents-research-methodology`).

AI & Automation 38 stars 13 forks Updated today MIT

Install

View on GitHub

Quality Score: 77/100

Stars 20%
53
Recency 20%
100
Frontmatter 20%
70
Documentation 15%
100
Issue Health 10%
50
License 10%
100
Description 5%
100

Skill Content

# AI Agents Research Frontier <!-- vendor-portability: contributor-facing knowledge pack for the rjmurillo/ai-agents repo itself; intentionally references upstream paths (.agents/, .claude/, scripts/, build/) because its audience is repo contributors, not plugin consumers (issue #2050) --> Open problems where this repository can advance the state of the art, ranked in owner-confirmed priority order. This skill tells you WHAT is worth working on and what "done" would look like. It does not teach experiment discipline (that is `ai-agents-research-methodology`) and it is not the portability battle plan (that is `ai-agents-portability-campaign`). Honesty contract for this document: every asset claim below was re-verified against the working tree on 2026-07-03. Anything labeled PROPOSED is not policy. Anything labeled UNVERIFIED could not be confirmed and must be checked before you build on it. ## Triggers - `research frontier` - `open problems` - `what should we research next` - `frontier programs` ## The Three Programs at a Glance | Rank | Program | Anchor artifacts | Status (as of 2026-07-03) | Falsifiable milestone (short form) | |------|---------|------------------|---------------------------|-------------------------------------| | 1 | Verified governance | ADR-069, `scripts/eval/eval-rule-activation.py` | ADR-069 is PROPOSED; eval tool exists, 7 rule scenario fixtures exist | Controlled eval shows gated-corpus sessions beat ungated on N scenarios with defensible stats...

Details

Author
rjmurillo
Repository
rjmurillo/ai-agents
Created
7 months ago
Last Updated
today
Language
Markdown
License
MIT

Integrates with

Similar Skills

Semantically similar based on skill content — not just same category