Relevance Does Not Imply Applicability: Experience Activation for Personal GUI Agents
cs.CV
Submitted: 2026-09-27
Updated: 2026-09-27
Terminology
Sources
- PIRA-Bench: A Transition from Reactive GUI Agents to GUI-based Proactive Intent Recommendation Agents
- KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation
- Graph2Eval: Automatic Multimodal Task Generation for Agents via Knowledge Graphs
- Evaluating Long-Context Reasoning in LLM-Based WebAgents
- Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History
- ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices
- ColorAgent: Building A Robust, Personalized, and Interactive OS Agent
- WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces
- Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization
- What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States
- PersonalAlign: Hierarchical Implicit Intent Alignment for Personalized GUI Agent with Long-Term User-Centric Records
- Proactive Agent Research Environment: Simulating Active Users to Evaluate Proactive Assistants
- PSPA-Bench: A Personalized Benchmark for Smartphone GUI Agent
- Executable Agentic Memory for GUI Agent
- Agent Workflow Memory
- FingerTip 20K: A Benchmark for Proactive and Personalized Mobile LLM Agents
- HippoCamp: Benchmarking Contextual Agents on Personal Computers
- OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks
- MementoGUI: Learning Agentic Multimodal Memory Control for Long-Horizon GUI Agents
- Hybrid Self-evolving Structured Memory for GUI Agents
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models