PEAR: Permutation-Equivariant Adaptive Routing Multi-Agent Debate
cs.AI, cs.CL, cs.LG, cs.MA, stat.ML
Submitted: 2026-05-26
Updated: 2026-08-31
Comments: Published in the 2026 Conference on Empirical Methods in Natural Language Processing (EMNLP) as a Main Conference paper
Code: https://github.com/EVIEHub/PEAR
License: http://creativecommons.org/licenses/by/4.0/
The gist: Multi-agent debate improves the reliability of large language models (LLMs) through iterative peer critiques.
Terminology
Abstract
Multi-agent debate improves the reliability of large language models (LLMs) through iterative peer critiques. However, fixed topologies often introduce persistent positional biases, amplify unreliable agents, and cause high sensitivity to role assignments. We introduce Permutation-Equivariant Adaptive Routing Multi-Agent Debate (PEAR), an inference-time train-free protocol that dynamically reconfigures communication roles and sparse topologies across consecutive debate rounds. By strategically switching agent-to-role assignments based on evolving agent states, PEAR prevents any agent from permanently occupying a privileged network position or distributes influence more evenly across the debate. We theoretically characterize PEAR as an equivariant sparse router: it preserves accuracy under agent relabeling while reducing routing complexity and improving generalization. Comprehensive empirical evaluations across four reasoning benchmarks and six diverse LLM backbones demonstrate PEAR significantly improves average accuracy over the strongest debate baselines. The code is available at https://github.com/EVIEHub/PEAR.
Sources
- Deep Think with Confidence
- The Llama 3 Herd of Models
- Measuring Mathematical Problem Solving With the MATH Dataset
- Learning Multi-Agent Communication from Graph Modeling Perspective
- Adaptive Graph Pruning for Multi-Agent Communication
- Training Verifiers to Solve Math Word Problems
- Enhancing Multi-Agent Debate System Performance via Confidence Expression
- GroupDebate: Enhancing the Efficiency of Multi-Agent Debate Using Group Discussion
- Qwen3 Technical Report
- Hear Both Sides: Efficient Multi-Agent Debate via Diversity-Aware Message Retention
- G-Designer: Architecting Multi-agent Communication Topologies via Graph Neural Networks
- Assemble Your Crew: Automatic Multi-agent Communication Topology Design via Autoregressive Graph Generation
- Demystifying Multi-Agent Debate: The Role of Confidence and Diversity
- Gemma 3 Technical Report
- Multi-Agent Debate with Memory Masking
- Self-Consistency Improves Chain of Thought Reasoning in Language Models
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection