Rethinking Legibility in Social Robot Hallway Navigation: Impact of Intent Representation and Human Distraction
summary
The gist
Legibility in social robot navigation is crucial for ensuring human safety and smooth coordination in dynamic, constrained environments where human attention can be divided.
In short
This research investigated how different ways of representing a robot's intent affect its ability to navigate safely and smoothly in hallways when humans are distracted. The study compared various intent models, finding that dynamic, interaction-level representations like Dynamic Passing Side Legibility were the most effective for coordination. Crucially, legible motion still improved objective coordination even when humans were divided in attention.
Key concepts
- Intent Representation
- This refers to how a robot communicates its plan or goal to a human. The study compared different methods, such as using the robot's final destination (goal-based) versus cues about which way it intends to pass (passing-side). The paper found that interaction-level cues are better than just stating a fixed destination.
- Legible Motion
- This is the quality of a robot's movement that makes its intentions clear to humans. The researchers tested several types, including goal-based and dynamic passing side legibility. They found that certain forms of legible motion lead to smoother human movements and better coordination during navigation.
- Divided Attention
- This occurs when a person is trying to do two things at once, like navigating a hallway while simultaneously listening to instructions. The study examined how distraction affects whether humans can correctly interpret the robot's signals. They found that even when distracted, legible motion still helps resolve conflicts.
- Dynamic Adaptation
- This involves a robot changing its behavior in real-time based on what it observes from the human. Dynamic Passing Side Legibility was superior because it automatically adjusted which way to pass based on predictions of human choice, showing that intent should evolve during interaction.
Terminology used across episodes
This episode discusses
- Rethinking Legibility in Social Robot Hallway Navigation: Impact of Intent Representation and Human Distraction · Paper Radio
- Responsibility and Engagement -- Evaluating Interactions in Social Robot Navigation
The paper
Rethinking Legibility in Social Robot Hallway Navigation: Impact of Intent Representation and Human Distraction · Read on arXiv
PRANAV GOYAL, ANDREW STRATTON, CHRISTOFOROS MAVROGIANNIS
University of Michigan at Ann Arbor
We focus on legible robot motion generation in social navigation settings. Legibility in human-robot interaction (HRI) is often described as the property of robot motion that enables an observer to confidently infer the robot's intent. While mature frameworks exist for generating legible motion in front of static observers, social robot navigation presents a new challenge: the robot must clearly convey its intent while ensuring human safety in dynamic pedestrian environments where human attention is often divided. With the goal of enabling robots to generate legible motion in dynamic and constrained spaces, we investigate how the choice of representation and the level of human attention shape navigation performance and human impressions. Focusing on the ubiquitous and demanding scenario of hallway navigation, we conduct two controlled user studies involving alternative legibility formulations implemented within a shared model predictive control framework. Study 1 (N = 45) investigates the role of intent representation, showing that passing-side legibility, particularly when adaptively updated, leads to smoother human motion and is perceived as more competent and less mentally and physically demanding than destination-based and non-legible baselines. Study 2 (N = 45) examines the effect of pedestrian attention, demonstrating that legible motion allows for smooth human motion even under distraction, even if this is not consistently reflected in subjective ratings. Together, these findings suggest that effective legible motion in social robot navigation benefits from interaction-level intent representations that support coordination, with some effects persisting even when human attention is divided. Code is available at https://github.com/fluentrobotics/Legible MPPI.
Transcript
Introduction to the show: ident: Robotics Radio. Generated commentary on the latest robotics and control papers.
Rosa: Today's paper: "Rethinking Legibility in Social Robot Hallway Navigation".
Dev: Legibility in social robot navigation is crucial for ensuring human safety and smooth coordination in dynamic, constrained environments where human attention can be divided.
Rosa: First, who's behind it and why it matters.
Paper summary: Rosa: Welcome everyone to our show today as we look at a really interesting paper on arXiv titled "Rethinking Legibility in Social Robot Hallway Navigation: Impact of Intent Representation and Human Distraction." This research dives into how robots should move in crowded spaces, especially when people are paying attention to other things. We'll discuss what the authors found regarding intent representation and how human distraction affects that legibility.
Dev: I’m ready for it, Rosa. Given my background in control systems, I’m particularly interested in whether these proposed intent representations translate into actually stable and low-latency movement when we put them into a Model Predictive Control framework, which is what this paper seems to be using.
Taro: From an autonomy standpoint, I want to hear about what happens when the environment gets messy; if the robot’s legibility cues don't hold up when pedestrians are distracted or the situation changes unexpectedly, how resilient is that motion strategy?
Rosa: Exactly, Taro. So basically, this paper looks at a major gap in existing research where we often only think about static observers and not dynamic ones in crowded hallways. The core thesis here is that how a robot communicates what it’s going to do—its intent—matters a lot more than just where it’s trying to go when humans are actually interacting with it.
Dev: So, the paper claims that moving beyond simple destination-based cues towards interaction-level coordination can be more effective in crowded settings, even when people aren't looking directly at the robot. That sounds like something we need to test on our hardware loop rates.
Taro: I’m interested in the specific representations they tested; are we talking about abstract concepts, or do they have concrete ways to encode that interaction-level coordination into the robot's actual trajectory planning? If it’s too abstract, it won't work when things go wrong.
Rosa: The authors compared several different ways to encode intent, ranging from simple goal-based legibility to more complex ideas like Social Momentum, and they found some really interesting trade-offs in hallway navigation. They specifically looked at how these different representations perform under conditions where human attention is divided during the interaction.
Dev: So, if I understand correctly, the paper suggests that some forms of intent representation are more robust than others when we can't rely on a person being perfectly focused on us? That has implications for our failure modes when we encounter unexpected human behavior or distraction.
Paper summary: Taro: If the paper shows that dynamic adaptation in intent—like selecting a passing side based on predicted human choice—works better than a fixed intent, that tells us we need more sophisticated real-time decision-making in our autonomy stacks to handle unpredictable social dynamics.
Rosa: That’s the big picture there. The whole point is to see how these different ways of encoding intent shape both how good the navigation actually is and what people feel when they observe it, especially when their attention is divided. It sets up a real challenge for designing robots that are not just safe, but also socially smooth in busy environments.
Dev: So the paper essentially argues that for social robot navigation in constrained settings like hallways, we need to focus on interaction-level cues rather than just destination-based ones, and that these cues should adapt based on what the human is actually doing or paying attention to.
Taro: And it suggests that even when those attention cues aren't perfectly clear subjectively, the objective measures of coordination still show a benefit from having a legible strategy in place. That persistence under distraction is something I think is important for real-world deployment because in reality, people are almost always distracted.
Rosa: Precisely, Taro. The authors’ conclusion emphasizes that adaptive strategies reinforce the legibility effect and that this coordination benefit continues even when subjective human impressions become less sensitive to the differences between strategies. This suggests we should build systems that can dynamically adjust their intent signaling based on observed human context during movement.
Dev: From a control engineering view, if the paper confirms that dynamic selection of passing sides, like in DPL or SM, yields better Human Average Acceleration metrics objectively, then our MPC framework needs to be able to incorporate those real-time predictions about human choice into its cost function for trajectory generation.
Taro: If the system needs to dynamically adapt its intent based on predicted human behavior during the interaction, that means our planning loop has to become much more tightly coupled with perception of the social context, not just the immediate geometric constraints of the hallway.
Rosa: So when we look at these results in "Rethinking Legibility in Social Robot Hallway Navigation: Impact of Intent Representation and Human Distraction," it’s clear that moving from simple destination-based planning to interaction-level intent representation is key for navigating crowded spaces safely and smoothly.
Dev: And the finding that adaptive strategies outperform fixed ones, even when people are distracted, points directly toward building more robust online estimation and adaptation capabilities into our navigation algorithms to handle those real-world human distractions effectively.
Paper summary: Taro: I think the biggest implication is that we need to design autonomy where the robot doesn't just follow a pre-set path but actively tries to maintain a legible social presence by constantly adjusting how it signals its intent based on what it senses about the human partner.
Rosa: That really puts things into perspective, Taro. The paper suggests that for these robots, legibility isn't just about moving smoothly in a vacuum; it’s about managing the perception of coordination in a messy hallway where attention shifts around.
Dev: I think we should be looking at how to implement those dynamic representations within our MPC framework to ensure low latency and reliable execution, even when the human input data is noisy due to distraction.
Taro: If we can figure out how to reliably estimate the human’s momentary focus or intent through sensory input, then that adaptive legibility becomes a powerful tool for achieving safer and more natural social interactions in dynamic environments.
Rosa: It sounds like this paper really pushes us toward integrating social context directly into the robot's motion planning decisions, which is exactly what we need to consider when we move these systems out of the controlled lab environment and into actual public spaces.
Dev: I’ll be checking how long these control loops can maintain that level of adaptability under sustained real-world operational stress; that’s where I see the biggest engineering hurdle for us right now.
Taro: That's a fair point, Dev. The future work they suggest about automatically balancing functional efficiency against legibility using an attention parameter lambda is precisely the kind of adaptation we need to explore for field deployment.
Rosa: So when we talk about the title "Rethinking Legibility in Social Robot Hallway Navigation: Impact of Intent Representation and Human Distraction," it really captures that this isn't just a technical tweak, but a fundamental rethinking of how robots should communicate their intentions in complex social settings.
Dev: It seems like the paper points us toward developing systems where intent is not just a static output from the planner, but something actively negotiated based on real-time interaction feedback and environmental awareness.
Taro: That means we’re not just building better path planners; we're building robots that are better at understanding and responding to the social state of their environment in real time.
Rosa: And that’s what makes this paper so compelling for all of us, showing how subtle shifts in intent encoding can lead to measurable improvements in both objective coordination and how people actually perceive the robot's behavior during a navigation task.
Conclusion: Rosa: So, to wrap up this discussion on "Rethinking Legibility in Social Robot Hallway Navigation: Impact of Intent Representation and Human Distraction," we've seen how changing how a robot signals its plan really affects both its performance and how people react when they're distracted.
Dev: Yeah, it’s clear the core focus here is moving beyond just where the robot is going to making sure that intention is actually legible to people interacting with it in busy hallways.
Taro: I think what stuck with me was how they showed that even when humans are distracted, those legibility cues still help resolve conflicts between robots or between robots and people.
Rosa: Exactly, Taro, and the authors found that using dynamic intent representations, like adapting the passing side based on predictions of human choice, works better than fixed ones.
Dev: From a control standpoint, if those adaptive strategies are producing lower Human Average Acceleration objectively in controlled tests while maintaining reasonable loop rates for MPC to handle them, that’s solid data.
Taro: But I wonder how resilient those systems really are when the world gets unexpectedly chaotic; does this hold up when the human partner suddenly changes their behavior drastically?
Rosa: That’s a fair question, Taro, and the authors themselves pointed out that while objective measures showed benefits under distraction, subjective impressions became less sensitive to algorithmic differences.
Dev: That means we need to figure out if those objective gains translate into reliable performance outside of controlled lab conditions where we can strictly script the interactions.
Taro: If we can transfer these concepts to real-world deployment, it suggests that robots will be much better at navigating crowded public spaces without needing a perfectly focused human partner.
Rosa: Exactly, and this work opens up a lot of ideas for how we design social robots that are not just efficient but also socially intuitive in messy environments.
Dev: We need to think about how to actually bake that dynamic adaptation into the MPC cost function so the system can handle those real-time adjustments smoothly without introducing latency issues.
Taro: So, the next step seems to be developing systems where robots can estimate human attention levels dynamically and adjust their legibility signals accordingly.
Rosa: That sounds like a really exciting direction for future research, Dev; we should definitely keep an eye on how those online estimation models evolve.
More episodes
- 2610.12154-Stochastic Distribution Network Reconfiguration under Load Uncertainty
- 2607.00148-3D Point World Models: Point Completion Enables More Accurate Dynamics Learning
- 2607.02403-ACID: Action Consistency via Inverse Dynamics for Planning with World Models
- 2510.26623-A Sliding-Window Filter for Online Continuous-Time Continuum Robot State Estimation
- 2406.13267-The Kinetics Observer: A Tightly Coupled Estimator for Legged Robots
- 2511.02147-Census-Based Population Autonomy For Distributed Robotic Teaming
- 2603.08260-Seed2Scale: A Self-Evolving Data Engine with Parallel Worlds Expansion for Scalable Robot Learning
- 2602.14032-RoboAug: One Annotation to Hundreds of Scenes via Region-Contrastive Data Augmentation for Robotic Manipulation
- 2602.15397-ActionCodec: What Makes for Good Action Tokenizers
- 2607.01819-Koopman operator theory: fundamentals, control, and applications