Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions
cs.CL, cs.AI, cs.CV, cs.LG
Submitted: 2026-08-31
Updated: 2026-08-31
Code: https://github.com/prismarinejs/mineflayer
Project page: https://junseokim0103.github.io/Lies-We-Can-See
Terminology
Sources
- AMONGAGENTS: Evaluating Large Language Models in the Interactive Text-Based Social Deduction Game
- The Traitors: Deception and Trust in Multi-Agent Language Model Simulations
- AI2-THOR: An Interactive 3D Environment for Visual AI
- Meta-Harness: End-to-End Optimization of Model Harnesses
- TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft
- Frontier Models are Capable of In-context Scheming
- Hoodwinked: Deception and Cooperation in a Text-Based Game for Language Models
- The MASK Benchmark: Disentangling Honesty From Accuracy in AI Systems
- CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
- Avalon's Game of Thoughts: Battle Against Deception through Recursive Contemplation
- Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf
- MineLand: Simulating Large-Scale Multi-Agent Interactions with Limited Multimodal Senses and Physical Needs
- VideoGameBench: Can Vision-Language Models complete popular video games?
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering