SkillVine: Agent Skill Evolution via Branching Exploration
cs.AI
Submitted: 2026-09-26
Updated: 2026-09-26
Code: https://github.com/EvolvingLMMs-Lab/SkillOpt-Lite
Terminology
Sources
- TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs
- SIMA 2: A Generalist Embodied Agent for Virtual Worlds
- Building Self-Evolving Agents via Experience-Driven Lifelong Learning: A Framework and Benchmark
- SearchQA: A New Q&A Dataset Augmented with Context from a Search Engine
- Go-Explore: a New Approach for Hard-Exploration Problems
- A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems
- rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
- LiveMathematicianBench: A Live Benchmark for Research-Level Mathematical Reasoning with Proof Sketches
- From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills
- Population Based Training of Neural Networks
- Tree Search for Language Model Agents
- The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
- Improve Mathematical Reasoning in Language Models by Automated Process Supervision
- Learning While Acting: A Skill-Enhanced Test-Time Co-Evolution Framework for Online Lifelong Learning Agents
- Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
- OfficeQA Pro: An Enterprise Benchmark for End-to-End Grounded Reasoning
- SkillOS: Learning Skill Curation for Self-Evolving Agents
- Anything2Skill: Compiling External Knowledge into Reusable Skills for Agents
- SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe
- More Skills, Worse Agents? Skill Shadowing Degrades Performance When Expanding Skill Libraries
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection