OdysSim: Building Foundation Models for Human Behavior Simulation
cs.CL, cs.AI, cs.LG
Submitted: 2026-06-12
Updated: 2026-09-07
Comments: 34 pages. Code: https://github.com/sunnweiwei/OdysSim ; Models and data: https://huggingface.co/collections/cmu-lti/odyssim ; v2: Corrected two bibliography entries
Code: https://github.com/sunnweiwei/OdysSim
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Out of One, Many: Using Language Models to Simulate Human Samples
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- Centaur: a foundation model of human cognition
- ConvoKit: A Toolkit for the Analysis of Conversations
- CaSiNo: A Corpus of Campsite Negotiation Dialogues for Automatic Negotiation Systems
- Persona Vectors: Monitoring and Controlling Character Traits in Language Models
- Chameleons in imagined conversations: A new approach to understanding coordination of linguistic style in dialogs
- SimulatorArena: Are User Simulators Reliable Proxies for Multi-Turn Evaluation of AI Assistants?
- TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
- Moral Stories: Situated Reasoning about Norms, Intents, Actions, and their Consequences
- Don't Stop Pretraining: Adapt Language Models to Domains and Tasks
- HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models
- Reinforcement Learning via Self-Distillation
- One Model, All Roles: Multi-Turn, Multi-Agent Self-Play Reinforcement Learning for Conversational Social Intelligence
- Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)
- FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions
- The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models
- PatientSim: A Persona-Driven Simulator for Realistic Doctor-Patient Interactions
- OpenAssistant Conversations -- Democratizing Large Language Model Alignment
- HumanLLM: Towards Personalized Understanding and Simulation of Human Nature
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering