MESSENGER: Memory-Enhanced Sequential Scene Flow Estimation via Autoregressive Next-Frame Forecasting
cs.CV
Submitted: 2026-10-07
Updated: 2026-10-07
Code: https://github.com/liujiuming123/Messenger
Terminology
Sources
- Binding Multiple Modalities via Multimodal Wasserstein Barycenter
- UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation
- TeFlow: Enabling Multi-frame Supervision for Self-Supervised Feed-forward Scene Flow Estimation
- EgoAction: Egocentric Action Composition with Reliability-Aware Temporal Fusion for the EPIC-KITCHENS Action Detection Challenge at CVPR 2026
- Mamba: Linear-Time Sequence Modeling with Selective State Spaces
- Point Mamba: A Novel Point Cloud Backbone Based on State Space Model with Octree-Based Ordering Strategy
- RegFormer++: An Efficient Large-Scale 3D LiDAR Point Registration Network with Projection-Aware 2D Transformer
- Argoverse 2: Next Generation Datasets for Self-Driving Perception and Forecasting
- Denoising Diffusion Implicit Models
- Wan: Open and Advanced Large-Scale Video Generative Models
- Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends
- WorldCoder-Bench: Benchmarking Physically Grounded 3D World Synthesis
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models