EVOKE: Eliciting World Knowledge in Agents for Transferable Decision-Making
cs.CL
Submitted: 2026-09-29
Updated: 2026-10-01
Terminology
Sources
- V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
- LoRA: Low-Rank Adaptation of Large Language Models
- AHEAD: Adaptive Hindsight with Environment-Augmented Distillation for Agentic RL
- Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models
- Policy and World Modeling Co-Training for Language Agents
- Self-Distilled Agentic Reinforcement Learning
- PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning
- Qwen2.5 Technical Report
- Contrastive Learning with Hard Negative Samples
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- ALFWorld: Aligning Text and Embodied Environments for Interactive Learning
- Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
- Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution
- MemWM: Memory-Augmented Text-Based World Model
- EnvRL: Learn from Environment Dynamics in Agentic Reinforcement Learning
- Qwen3 Technical Report
- Self-Distilled RLVR
- OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning
- Reinforcement World Model Learning for LLM-based Agents
- Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering