AffordDrive3D: Affordance-Aware World-Action Modeling with Spatial Understanding
cs.CV, cs.AI
Submitted: 2026-10-08
Updated: 2026-10-08
Terminology
Sources
- Qwen3-VL Technical Report
- OWMDrive: Causality-Aware End-to-End Autonomous Driving via 4D Occupancy World Model
- DriveFuture: Future-Aware Latent World Models for Autonomous Driving
- Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation
- Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer
- Generalized Trajectory Scoring for End-to-end Multimodal Planning
- UniFuture: A 4D Driving World Model for Future Generation and Perception
- Depth Anything 3: Recovering the Visual Space from Any Views
- DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
- GeoWAM: Visual Geometry World Action Models for Autonomous Driving
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving
- Wan: Open and Advanced Large-Scale Video Generative Models
- Learning Vision-Language-Action World Models for Autonomous Driving
- Afford-VLA: Action-Aligned Visual Planning via Internalized Affordance
- WA-JEPA: Rethinking the Video JEPA Paradigm for World-Action Modeling in Autonomous Driving
- HAD: Combining Hierarchical Diffusion with Metric-Decoupled RL for End-to-End Driving
- AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models