DEALS: Decentralized Expertise-Aware Load Serving for Multi-Agent LLM Systems
cs.LG, cs.AI, cs.MA
Submitted: 2026-09-27
Updated: 2026-09-27
Terminology
Sources
- RouteBalance: Fused Model Routing and Load Balancing for Heterogeneous LLM Serving
- Throughput-Optimal Scheduling Algorithms for LLM Inference and AI Agents
- Improving Factuality and Reasoning in Language Models through Multiagent Debate
- The Llama 3 Herd of Models
- Large Language Model based Multi-Agents: A Survey of Progress and Challenges
- Self-Evolving Multi-Agent Systems via Decentralized Memory
- Measuring Mathematical Problem Solving With the MATH Dataset
- More Agents Is All You Need
- Markets, Not Planners: Decentralized Orchestration of LLM Agents with Private Information
- A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
- MorphAgent: Empowering Agents through Self-Evolving Profiles and Decentralized Collaboration
- Autellix: An Efficient Serving Engine for LLM Agents as General Programs
- Chimera: Latency- and Performance-Aware Multi-agent Serving for Heterogeneous LLMs
- Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems
- AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation
- Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
- ReAct: Synergizing Reasoning and Acting in Language Models
- The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
- Augmented Backpressure for Decentralized Management of Agentic Networks
- Language Agents as Optimizable Graphs
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks