KorNAT: LLM Alignment Benchmark for Korean Social Values and Common Knowledge
Jiyoung Lee, Minwoo Kim, Seungho Kim, Junghwan Kim, Seunghyun Won, Hwaran Lee, Edward Choi
cs.CL
Submitted: 2026-08-19
Updated: 2026-08-21
Comments: Accepted at ACL 2024 Findings (35 pages, 7 figures, 16 tables)
Code: https://github.com/jiyounglee-0523/KorNAT
License: http://creativecommons.org/licenses/by/4.0/
The gist: For Large Language Models (LLMs) to be effectively deployed in a specific country, they must possess an understanding of the nation's culture and basic knowledge.
Terminology
Abstract
For Large Language Models (LLMs) to be effectively deployed in a specific country, they must possess an understanding of the nation's culture and basic knowledge. To this end, we introduce National Alignment, which measures an alignment between an LLM and a targeted country from two aspects: social value alignment and common knowledge alignment. Social value alignment evaluates how well the model understands nation-specific social values, while common knowledge alignment examines how well the model captures basic knowledge related to the nation. We constructed KorNAT, the first benchmark that measures national alignment with South Korea. For the social value dataset, we obtained ground truth labels from a large-scale survey involving 6,174 unique Korean participants. For the common knowledge dataset, we constructed samples based on Korean textbooks and GED reference materials. KorNAT contains 4K and 6K multiple-choice questions for social value and common knowledge, respectively. Our dataset creation process is meticulously designed and based on statistical sampling theory and was refined through multiple rounds of human review. The experiment results of seven LLMs reveal that only a few models met our reference score, indicating a potential for further enhancement. KorNAT has received government approval after passing an assessment conducted by a government-affiliated organization dedicated to evaluating dataset quality. Samples and detailed evaluation protocols of our dataset can be found in https://huggingface.co/datasets/jiyounglee0523/KorNAT and https://github.com/jiyounglee-0523/KorNAT.
Sources
- PaLM 2 Technical Report
- A General Language Assistant as a Laboratory for Alignment
- Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
- Disentangling Perceptions of Offensiveness: Cultural and Moral Correlates
- Towards Measuring the Representation of Subjective Global Opinions in Language Models
- CALM : A Multi-task Benchmark for Comprehensive Assessment of Language Model Bias
- CBBQ: A Chinese Bias Benchmark Dataset Curated with Human-AI Collaboration for Large Language Models
- What Makes Good In-Context Examples for GPT-$3$?
- KoBBQ: Korean Bias Benchmark for Question Answering
- Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering
- Alignment of Language Agents
- Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?
- RACE: Large-scale ReAding Comprehension Dataset From Examinations
- KoSBi: A Dataset for Mitigating Social Bias Risks Towards Safer Large Language Model Application
- Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
- The Tail Wagging the Dog: Dataset Construction Biases of Social Bias Benchmarks
- Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
- Gemini: A Family of Highly Capable Multimodal Models
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- Learning from the Worst: Dynamically Generated Datasets to Improve Online Hate Detection
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering