Which Self-Improvements Should We Trust? Reliable Self-Improvement When Agents Reuse Their Benchmarks
stat.ML, cs.AI, cs.LG, stat.ME
Submitted: 2026-09-27
Updated: 2026-09-29
Terminology
Sources
- What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents
- Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops
- Optimal Inference After Model Selection
- Self-Authored Verification Is Unreliable in Heuristic Self-Improving Agents
- The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
- Categorizing Variants of Goodhart's Law
- Signed Compression Progress on a Sealed Audit is Goodhart-Resistant
- AlphaEvolve: A coding agent for scientific and algorithmic discovery
- Self-Evolving Agents with Anytime-Valid Certificates
- PACE: Anytime-Valid Acceptance Tests for Self-Evolving Agents
- SGM: A Statistical Godel Machine for Risk-Controlled Recursive Self-Modification
- Towards Reliable LLM Evaluation: Correcting the Winner's Curse in Adaptive Benchmarking
- Qwen2.5 Technical Report
Related papers
- Behavior of prediction performance metrics with rare events
- Optimal Estimation of Generic Dynamics by Path-Dependent Neural Jump ODEs
- A Posterior-Dynamics Framework for Imaging Inverse Problems with Pretrained Diffusion Priors
- One Permutation Is All You Need: Fast, Deterministic Feature Importance and Model Stress-Testing
- Online Conformal Prediction for Non-Exchangeable Panel Data
- Deep Time-Series Forecasting in 10 Years: A Survey