Mind2Dialogue: Training Human-Aware Language Models by Simulating User Mental States
cs.CL, cs.LG
Submitted: 2026-09-14
Updated: 2026-09-14
Comments: 40 pages, 10 figures, 11 tables. Project page: https://wannabeyourfriend.github.io/mind2dialogue/
Project page: https://wannabeyourfriend.github.io/mind2dialogue
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
- User Simulation with Large Language Models for Evaluating Task-Oriented Dialogue
- Scaling Synthetic Data Creation with 1,000,000,000 Personas
- The Llama 3 Herd of Models
- Distilling the Knowledge in a Neural Network
- PersonaMem-v2: Towards Personalized Intelligence via Learning Implicit User Personas and Agentic Memory
- HorizonBench: Long-Horizon Personalization with Evolving Preferences
- Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation
- MemAudit: Auditing Long-Term Agent Memory via Hidden User-State Recovery
- Synthetic Interaction Data for Scalable Personalization in Large Language Models
- Olmo 3
- UserRL: Training Interactive User-Centric Agent via Reinforcement Learning
- Qwen2.5 Technical Report
- Taxonomy of User Needs and Actions
- OpenAI GPT-5 System Card
- Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks
- DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas
- HumanLM: Simulating Users with State Alignment Beats Response Imitation
- OdysSim: Building Foundation Models for Human Behavior Simulation
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering