A research framework for principled agent self-improvement under frozen evaluators and declared mutation boundaries, recording verifiable lineage to make it reproducible and auditable.
-
Updated
Sep 5, 2026 - Python
A research framework for principled agent self-improvement under frozen evaluators and declared mutation boundaries, recording verifiable lineage to make it reproducible and auditable.
AI4AI Survey: can AI reliably improve AI? 223 papers on long-horizon agents, benchmarks, harness design, and recursive self-improvement · updated weekly
AI Systems as Scientific Instruments for Understanding the Mechanisms of Intelligence
[EMNLP26 Findings] Official repository for DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning
To associate your repository with the ai4ai topic, visit your repo's landing page and select "manage topics."