LEGO: A Lifting-Free Approach for Exocentric-to-Egocentric Video Generation
cs.CV
Submitted: 2026-10-08
Updated: 2026-10-08
Code: https://github.com/suhwan-cho/lego
Terminology
Sources
- Qwen2.5-VL Technical Report
- Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
- E$^3$C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control
- ViPE: Video Pose Engine for 3D Geometric Perception
- FreqForcing: Autoregressive Long Video Generation via Spectral Self-Anchoring
- Intention-driven Ego-to-Exo Video Generation
- DINOv3
- Towards Accurate Generative Models of Video: A New Metric & Challenges
- Wan: Open and Advanced Large-Scale Video Generative Models
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models