Which Skill to Distill? SGUID: Selecting a Compact Skill Bank for Model-Skill Co-Evolution
cs.CL
Submitted: 2026-10-08
Updated: 2026-10-08
Terminology
Sources
- EvoSkill: Automated Skill Discovery for Multi-Agent Systems
- $\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment
- Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs
- Metacognitive Reuse: Turning Recurring LLM Reasoning Into Concise Behaviors
- daVinci-kernel: Co-Evolving Skill Selection, Summarization, and Utilization via RL for GPU Kernel Optimization
- Mathematical exploration and discovery at scale
- SWE-Skills-Bench: Do Agent Skills Actually Help in Real-World Software Engineering?
- LoRA: Low-Rank Adaptation of Large Language Models
- Skill-Conditioned Gated Self-Distillation for LLM Reasoning
- Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills
- Reinforcement Learning via Self-Distillation
- Tmax: A simple recipe for terminal agents
- XSkill: Continual Learning from Experience and Skills in Multimodal Agents
- SkillFlow: Scalable and Efficient Agent Skill Retrieval System
- SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
- DeepSeek-V3 Technical Report
- Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
- Humanity's Last Exam
- Future of Work with AI Agents: Auditing Automation and Augmentation Potential across the U.S. Workforce
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering