Do Large Language Models Possess a Theory of Mind? A Comparative Evaluation Using the Strange Stories Paradigm
cs.CL, cs.AI
Submitted: 2026-02-20
Updated: 2026-08-29
Terminology
Sources
- Evaluating Large Language Models in Theory of Mind Tasks
- ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
- EmoBench: Evaluating the Emotional Intelligence of Large Language Models
- SocialIQA: Commonsense Reasoning about Social Interactions
- Theory of Mind in Large Language Models: Examining Performance of 11 State-of-the-Art models vs. Children Aged 7-10 on Advanced Tests
- Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
- Mind Your Theory: Theory of Mind Goes Deeper Than Reasoning
- CogLM: Tracking Cognitive Development of Large Language Models
- Back to the Future: Towards Explainable Temporal Reasoning with Large Language Models
- A Survey of Large Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering