Exploring Collaboration between a language and a non-language agent
cs.CL, cs.AI
Submitted: 2026-08-31
Updated: 2026-09-02
Code: https://github.com/lightvector/KataGo
Project page: https://behavior-in-the-wild.github.io/llamia.html
Terminology
Sources
- RT-1: Robotics Transformer for Real-World Control at Scale
- RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
- PaLM-E: An Embodied Multimodal Language Model
- Training Large Language Models to Reason in a Continuous Latent Space
- MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework
- Evidence of Learned Look-Ahead in a Chess-Playing Neural Network
- Bridging the Gap between Expert and Language Models: Concept-guided Chess Commentary Generation and Evaluation
- LLM CHESS: Benchmarking Reasoning and Instruction-Following in LLMs through Chess
- Improving Chess Commentaries by Combining Language Models with Symbolic Reasoning Engines
- Tracing the Thought of a Grandmaster-level Chess-Playing Transformer
- Visual Instruction Tuning
- G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment
- Mastering Chess with a Transformer Model
- Toolformer: Language Models Can Teach Themselves to Use Tools
- Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning
- Learning Multiagent Communication with Backpropagation
- Enhancing Latent Computation in Transformers with Latent Tokens
- Multi-Agent Collaboration Mechanisms: A Survey of LLMs
- Latent Collaboration in Multi-Agent Systems
- Accelerating Self-Play Learning in Go
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering