Evolving Excellence: Automated Optimization of LLM-based Agents
cs.SE, cs.AI
Submitted: 2025-12-09
Updated: 2026-09-03
Code: https://github.com/pppyb/mini-swe-agent
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
- AFlow: Automating Agentic Workflow Generation
- Why Do Multi-Agent LLM Systems Fail?
- Training Verifiers to Solve Math Word Problems
- Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution
- Language Models for Code Optimization: Survey, Challenges and Future Directions
- Automated Design of Agentic Systems
- ALE-Bench: A Benchmark for Long-Horizon Objective-Driven Algorithm Engineering
- ShinkaEvolve: Towards Open-Ended And Sample-Efficient Program Evolution
- AlphaEvolve: A coding agent for scientific and algorithmic discovery
- Towards Scientific Intelligence: A Survey of LLM-based Scientific Agents
- Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering
- ReAct: Synergizing Reasoning and Acting in Language Models
- Large Language Models Are Human-Level Prompt Engineers
Related papers
- Falsification-Based Verification of LLM-Generated Optimization Models: Sound Test Batteries and Their Detection Limits
- GitSkills: A Dataset of Agent Skills on GitHub
- SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
- PackMonitor: Enabling Zero Package Hallucinations Through Decoding-Time Monitoring
- IntentCoding: Amplifying User Intent in Code Generation
- Incentives and Outcomes in Bug Bounties