HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos
cs.RO, cs.AI, cs.CV, cs.LG
Submitted: 2026-05-24
Updated: 2026-09-14
Comments: Project page: https://humanego-ai.github.io
Code: https://github.com/TX-Leo/HumanEgo
Project page: https://humanego-ai.github.io
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
- ALOHA 2: An Enhanced Low-Cost Hardware for Bimanual Teleoperation
- Project Aria: A New Tool for Egocentric Multi-Modal AI Research
- ImMimic: Cross-Domain Imitation from Human Videos via Mapping and Interpolation
- EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
- EgoScale: Scaling Dexterous Manipulation with Diverse Egocentric Human Data
- EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World
- Masquerade: Learning from In-the-wild Human Videos using Data-Editing
- EmbodiSwap for Zero-Shot Robot Imitation Learning
- EgoZero: Robot Learning from Smart Glasses
- H2R: A Human-to-Robot Data Augmentation for Robot Pre-training from Videos
- Any-point Trajectory Modeling for Policy Learning
- GenFlowRL: Shaping Rewards with Generative Object-Centric Flow in Visual Reinforcement Learning
- NovaFlow: Zero-Shot Manipulation via Actionable Flow from Generated Videos
- Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations
- Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking
- UniDex: A Robot Foundation Suite for Universal Dexterous Hand Control from Egocentric Human Videos
- TraceGen: World Modeling in 3D Trace Space Enables Learning from Cross-Embodiment Videos
- Dexterity from Smart Lenses: Multi-Fingered Robot Manipulation with In-the-Wild Human Demonstrations
- AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation
Related papers
- FMT x: An Efficient and Asymptotically Optimal Extension of the Fast Marching Tree for Dynamic Replanning
- MPCFormer: A physics-informed data-driven approach for explainable socially-aware autonomous driving
- RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
- HRDexDB: A 4D Dexterous Grasping Dataset Across Human and Multiple Robot Embodiments
- APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
- Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving