An Informational Curse of Horizon in Goal-Conditioned Policy Learning
cs.LG, cs.AI, cs.RO
Submitted: 2026-10-07
Updated: 2026-10-07
Terminology
Sources
- Learning Universal Policies via Text-Guided Video Generation
- Three Steps at a Time: Learning Representations from Action Sequences in Contrastive RL
- Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review
- Causal World Modeling for Robot Control
- Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio
- Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
- ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities
- Act2Goal: From World Model To General Goal-conditioned Policy
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks