FIVE-VLA: Fast and EffectIVE Autonomous Driving with Recurrent Action Memory
cs.CV, cs.RO
Submitted: 2026-09-16
Updated: 2026-09-16
Code: https://github.com/OpenDriveLab/DriveLM
Terminology
Sources
- Qwen Technical Report
- GPT-4 Technical Report
- Gemini: A Family of Highly Capable Multimodal Models
- Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail
- Towards Efficient and Effective Multi-Camera Encoding for End-to-End Driving
- ARC Prize 2024: Technical Report
- Hierarchical Reasoning Model
- Less is More: Recursive Reasoning with Tiny Networks
- Qwen2 Technical Report
- MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications
- Layer Normalization
- Fail2Drive: Benchmarking Closed-Loop Driving Generalization
- Qwen3 Technical Report
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models