Fewer Tokens, Better Action: GPT-6 Astra Robot Agents with 14% Higher Success Rate but 65% Fewer Tokens
cs.CV, cs.RO
Submitted: 2026-10-01
Updated: 2026-10-01
Code: https://github.com/DAGroup-PKU/PyRUA-Lean
Project page: https://dagroup-pku.github.io/PyRUA-Lean
Terminology
Sources
- SAM 3: Segment Anything with Concepts
- RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
- RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies
- Show-Harness: Just a VLM Agent Can Play Robots
- RLDX-1 Technical Report
- ASPIRE: Agentic /Skills Discovery for Robotics
- $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
- Voyager: An Open-Ended Embodied Agent with Large Language Models
- A Pragmatic VLA Foundation Model
- Harness VLA: Steering Frozen VLAs into Reliable Manipulation Primitives via Memory-Guided Agents
- LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models