MOSAIC: A Universal Agent-Level Interface for Cross-Paradigm Agent Mixing and Human-AI Collaboration
cs.LG, cs.AI
Submitted: 2026-03-01
Updated: 2026-09-10
Comments: 4 pages, 2 figures
Code: https://github.com/Abdulhamid97Mousa/MOSAIC
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- An Introduction to Vision-Language Modeling
- Game Reasoning Arena: A Framework and Benchmark for Assessing Reasoning Capabilities of Large Language Models via Game Play
- GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents
- TextArena
- Acme: A Research Framework for Distributed Reinforcement Learning
- lmgame-Bench: How Good are LLMs at Playing Games?
- OpenRL: A Unified Reinforcement Learning Framework
- LLM CHESS: Benchmarking Reasoning and Instruction-Following in LLMs through Chess
- BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
- TextAtari: 100K Frames Game Playing with Language Agents
- LLM-PySC2: Starcraft II learning environment for Large Language Models
- XuanCe: A Comprehensive and Unified Deep Reinforcement Learning Library
- BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games
- PettingZoo: Gym for Multi-Agent Reinforcement Learning
- Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard
- Atari-GPT: Benchmarking Multimodal Large Language Models as Low-Level Policies in Atari Games
- AgentGym: Evolving Large Language Model-based Agents across Diverse Environments
- CREW: Facilitating Human-AI Teaming Research
- MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks