Ecdysis: Efficient and Effective Training of Runtime Harnesses for LLM Agents
cs.SE, cs.AI
Submitted: 2026-09-10
Updated: 2026-09-20
Code: https://github.com/cuiyu-ai/Ecdysis
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Recursive Harness Self-Improvement
- From Failed Trajectories to Reliable LLM Agents: Diagnosing and Repairing Harness Flaws
- Harness-R1: Learning to Edit Executable Runtime Harnesses from Agent Failure Trajectories
- HarnessEvolve: Learning from Reference Trajectories for Reliable Agent Self-Evolution
- Self-Harness: Harnesses That Improve Themselves
- OpenForgeRL: Train Harness-native Agents in Any Environment
- MemoHarness: Agent Harnesses That Learn from Experience
- Verify Smarter, Evolve Further: Efficient Harness Evolution through Behavior-Aware Verification
- Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM Agents
- Qwen3 Technical Report
- DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
- What Matters in On-Policy Distillation? A Perspective on Data Efficiency and Data Selection
Related papers
- Falsification-Based Verification of LLM-Generated Optimization Models: Sound Test Batteries and Their Detection Limits
- GitSkills: A Dataset of Agent Skills on GitHub
- SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
- PackMonitor: Enabling Zero Package Hallucinations Through Decoding-Time Monitoring
- IntentCoding: Amplifying User Intent in Code Generation
- Incentives and Outcomes in Bug Bounties