ViLoMan: Learning Visual-Proprioceptive Whole-Body Loco-Manipulation Skills for Humanoid Robots
cs.RO
Submitted: 2026-09-16
Updated: 2026-09-16
Terminology
Sources
- HDMI: Learning Interactive Humanoid Whole-Body Control from Human Videos
- ResMimic: From General Motion Tracking to Humanoid Whole-body Loco-Manipulation via Residual Learning
- SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework
- VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation
- Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors
- OmniContact: Chaining Meta-Skills via Contact Flow for Generalizable Humanoid Loco-Manipulation
- VIRAL: Visual Sim-to-Real at Scale for Humanoid Loco-Manipulation
- Opening the Sim-to-Real Door for Humanoid Pixel-to-Action Policy Transfer
- StageACT: Stage-Conditioned Imitation for Robust Humanoid Door Opening
- OASIS: From Simulation Data Collection to Real-World Humanoid Loco-Manipulation
- VAIC: Vision-Guided Humanoid Agile Object Interaction Control via Decoupled Commands
- Retargeting Matters: General Motion Retargeting for Humanoid Motion Tracking
- OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipulation and Scene Interaction
- GRAIL: Generating Humanoid Loco-Manipulation from 3D Assets and Video Priors
- ExoActor: Exocentric Video Generation as Generalizable Interactive Humanoid Control
- GentleHumanoid: Learning Upper-body Compliance for Contact-rich Human and Object Interaction
- SceneBot: Contact-Prompted General Humanoid Whole Body Tracking with Scene-Interaction
- Perceptive Humanoid Parkour: Chaining Dynamic Human Skills via Motion Matching
- Kimodo: Scaling Controllable Human Motion Generation
- DoorGym: A Scalable Door Opening Environment And Baseline Agent
Related papers
- FMT x: An Efficient and Asymptotically Optimal Extension of the Fast Marching Tree for Dynamic Replanning
- MPCFormer: A physics-informed data-driven approach for explainable socially-aware autonomous driving
- RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
- HRDexDB: A 4D Dexterous Grasping Dataset Across Human and Multiple Robot Embodiments
- APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
- Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving