ActiveArena: Benchmarking and Understanding Active Perception in Robotic Manipulation
cs.RO, cs.AI
Submitted: 2026-09-21
Updated: 2026-09-23
Comments: 43 pages. Project page: https://leeibo.github.io/ActiveArena
Project page: https://leeibo.github.io/ActiveArena
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Qwen3-VL Technical Report
- $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
- RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
- RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies
- RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design
- EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents
- StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing
- RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies
- SAM2Act: Integrating Visual Foundation Model with A Memory Architecture for Robotic Manipulation
- LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models
- Towards Human-level Intelligence via Human-like Whole-Body Manipulation
- ManiSkill2: A Unified Benchmark for Generalizable Manipulation Skills
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics
- Towards Exploratory and Focused Manipulation with Bimanual Active Perception: A New Problem, Benchmark and Strategy
- ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop
- VIMA: General Robot Manipulation with Multimodal Prompts
- OpenVLA: An Open-Source Vision-Language-Action Model
- Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
- RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark
- BEHAVIOR-1K: A Human-Centered, Embodied AI Benchmark with 1,000 Everyday Activities and Realistic Simulation
Related papers
- FMT x: An Efficient and Asymptotically Optimal Extension of the Fast Marching Tree for Dynamic Replanning
- MPCFormer: A physics-informed data-driven approach for explainable socially-aware autonomous driving
- RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
- HRDexDB: A 4D Dexterous Grasping Dataset Across Human and Multiple Robot Embodiments
- APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
- Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving