Meta n: Recursive Self-Improvement through Emergent Depth
cs.AI, cs.CL, cs.SY, eess.SY
Submitted: 2026-08-25
Updated: 2026-08-25
Code: https://github.com/minnesotanlp/meta-n
License: http://creativecommons.org/licenses/by-nc-nd/4.0/
The gist: Self-improving LLM agents refine answers, not the process that produces those answers.
Terminology
Abstract
Self-improving LLM agents refine answers, not the process that produces those answers. Systems that add a meta-level hold that level fixed, and those that edit themselves must leave part of their own editing machinery untouched to stay stable, capping the meta-depth they realize at roughly two. We present Meta n, which keeps the meta-operation fixed and recurses on its input instead. That operation, Ω, is applied repeatedly to its own products, reading the traces of the solver stack below together with the code that produced them, then writing the next layer as a strategic pre-process and a library of callable helpers. Because Ω never changes, it cannot destabilize the system, and because its input strictly grows, each layer reasons from a higher vantage than the last. Depth is set by convergence rather than fixed in advance, and an evolutionary archive searches over layer chains. Across two backbones, Meta n outperforms prior self-improving agents on all eight benchmark families. The sharpest case is ARC-AGI-2, built to resist skill memorization, where it alone scores above zero. Ablations indicate that most of the gain from recursion comes from the conditioning each layer passes to the next, and distinct layer roles emerge with depth although no prompt prescribes them. Code available at https://github.com/minnesotanlp/meta-n
Sources
- EvoX: Meta-Evolution for Automated Discovery
- AlphaEvolve: A coding agent for scientific and algorithmic discovery
- Goedel Machines: Self-Referential Universal Problem Solvers Making Provably Optimal Self-Improvements
- G\"odel Agent: A Self-Referential Agent Framework for Recursive Self-Improvement
- Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection