DualManip: Agentic Dynamic Manipulation via Dual-Path Semantic Reasoning and Geometric Adaptation
cs.RO
Submitted: 2026-09-25
Updated: 2026-09-29
Project page: https://lichengxi1.github.io/Dualmanip
Terminology
Sources
- Towards Generalizable Robotic Manipulation in Dynamic Environments
- DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
- DynamicWAM: Dual-Path Motion Conditioning for World-Action Models in Dynamic Manipulation
- DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation
- Sparse Meets Dense: Correspondence Guided Robotic Manipulation with Rigid-Deformable Interactions
- Vi-TacMan: Articulated Object Manipulation via Vision and Touch
- DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation
- VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
- GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation
- TiPToP: A Modular Open-Vocabulary Robot Manipulation System That Plans
- UniManip: General-Purpose Zero-Shot Robotic Manipulation with Agentic Operational Graph
- ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
- RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
- OpenVLA: An Open-Source Vision-Language-Action Model
- $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
- $\tau_0$-WM: A Unified Video-Action World Model for Robotic Manipulation
- OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation
- ABot-M0.5: Unified Mobility-and-Manipulation World Action Model
- LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence
- Executable Analytic Concepts as the Missing Link Between VLM Insight and Precise Manipulation
Related papers
- FMT x: An Efficient and Asymptotically Optimal Extension of the Fast Marching Tree for Dynamic Replanning
- MPCFormer: A physics-informed data-driven approach for explainable socially-aware autonomous driving
- RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
- HRDexDB: A 4D Dexterous Grasping Dataset Across Human and Multiple Robot Embodiments
- APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
- Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving