InterEvolve: Test-Time Evolution of Reward Programs for Humanoid Loco-Manipulation
cs.RO, cs.CV, cs.GR
Submitted: 2026-10-01
Updated: 2026-10-04
Project page: https://sirui-xu.github.io/InterEvolve
Terminology
Sources
- Exploration and Online Transfer with Behavioral Foundation Models
- Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
- TEXEDO : Test Time Scaling for Controller-aware Language-conditioned Humanoid Motion Generation
- Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors
- GMT: General Motion Tracking for Humanoid Whole-Body Control
- Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration
- RHO: Your Coding Agent is Secretly a Roboticist
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation
- ExBody2: Advanced Expressive Humanoid Whole-Body Control
- DreamControl: Human-Inspired Whole-Body Humanoid Control for Scene Interaction via Guided Diffusion
- RDA: Reward Design Agent for Reinforcement Learning
- Learning to Learn Faster from Human Feedback with Language Model Predictive Control
- ASPIRE: Agentic /Skills Discovery for Robotics
- Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning
- AlphaEvolve: A coding agent for scientific and algorithmic discovery
- Cybo-Waiter: A Physical Agentic Framework for Humanoid Whole-Body Locomotion-Manipulation
- Optimistic Task Inference for Behavior Foundation Models
- RGB: RL Guided Whole-Body MPPI for Humanoid Control
- Fast Adaptation with Behavioral Foundation Models
Related papers
- FMT x: An Efficient and Asymptotically Optimal Extension of the Fast Marching Tree for Dynamic Replanning
- MPCFormer: A physics-informed data-driven approach for explainable socially-aware autonomous driving
- RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
- HRDexDB: A 4D Dexterous Grasping Dataset Across Human and Multiple Robot Embodiments
- APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
- Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving