CultureConverse: A Multilingual Multi-turn Simulation Harness for Culturally Grounded Assistance in East and Southeast Asia
cs.CL, cs.CY
Submitted: 2026-08-28
Updated: 2026-09-29
Code: https://github.com/Social-AI-Studio/CultureConverse
Terminology
Sources
- LoraxBench: A Multitask, Multilingual Benchmark Suite for 20 Indonesian Languages
- A General Language Assistant as a Laboratory for Alignment
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
- Bias and Fairness in Large Language Models: A Survey
- Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following
- Values in the Wild: Discovering and Analyzing Values in Real-World Language Model Interactions
- AlignCultura: Towards Culturally Aligned Large Language Models?
- The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models
- Quantifying the Persona Effect in LLM Simulations
- CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries
- Cultural Alignment in Large Language Models Using Soft Prompt Tuning
- A Survey on Bias and Fairness in Machine Learning
- Do Large Language Models Understand Morality Across Cultures?
- SEA-LION: Southeast Asian Languages in One Network
- LiveCultureBench: a Multi-Agent, Multi-Cultural Benchmark for Large Language Models in Dynamic Social Simulations
- Typhoon: Thai Large Language Models
- SEA-HELM: Southeast Asian Holistic Evaluation of Language Models
- Can Persona-Prompted LLMs Emulate Subgroup Values? An Empirical Analysis of Generalisability and Fairness in Cultural Alignment
- Value Compass Benchmarks: A Platform for Fundamental and Validated Evaluation of LLMs Values
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering