When Personality Meets Quantization: A Layer-wise MBTI Analysis of Quantized LLMs
cs.CL
Submitted: 2026-08-26
Updated: 2026-09-15
Code: https://github.com/AutoGPTQ/AutoGPTQ
License: http://creativecommons.org/licenses/by/4.0/
The gist: Personality is increasingly important in large language models (LLMs), as it shapes users' trust, engagement, and emotional experiences.
Terminology
Abstract
Personality is increasingly important in large language models (LLMs), as it shapes users' trust, engagement, and emotional experiences. While the Myers--Briggs Type Indicator (MBTI) has emerged as a common framework for assessing LLMs' personality, existing studies focus primarily on full-precision models and evaluate only final outputs. They overlook the widespread deployment of quantized LLMs requiring low memory footprints, whose personality traits remain underexplored. In this work, we present a systematic MBTI analysis of open-source LLMs across multiple precisions, including mainstream 4-bit methods (GPTQ, AWQ) and extreme 2-bit settings (AQLM variants). Beyond output-level evaluation, we examine how personality emerges across layers through option-level entropy and confidence-gap dynamics, and introduce Uncertainty-Amplified Layer Decoding (UALD) to study decoding-induced personality drift at inference time. Our results reveal a key insight: LLMs' personality is not a static property, but an emergent, layer-dependent decision process sensitive to quantization, prompting, and decoding. Specifically, we find that (1) ENFJ remains dominant across model families and precisions; (2) 4-bit quantization largely preserves coarse personality structure, while 2-bit quantization disrupts fine-grained prompt consistency and cross-precision agreement; (3) personality decisions emerges in upper layers, following substantial ambiguity in early layers; and (4) inference decoding can shift personality, while personality-aligned conditioning improves robustness. These findings provide a new perspective on the behavioral reliability of quantized LLMs and highlight the importance of considering internal dynamics and inference strategies in personality-sensitive chatbot applications.
Sources
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- HarmLevelBench: Evaluating Harm-Level Compliance and the Impact of Quantization on Model Alignment
- Identifying and Manipulating the Personality Traits of Language Models
- INT2.1: Towards Fine-Tunable Quantized Large Language Models with Error Correction through Low-Rank Adaptation
- Towards Understanding and Improving Refusal in Compressed Models via Mechanistic Interpretability
- DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models
- Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
- BitDistiller: Unleashing the Potential of Sub-4-Bit LLMs via Self-Distillation
- The Llama 3 Herd of Models
- Exploiting LLM Quantization
- Extreme Compression of Large Language Models via Additive Quantization
- GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
- LQ-LoRA: Low-rank Plus Quantized Matrix Decomposition for Efficient Language Model Finetuning
- LoRA+: Efficient Low Rank Adaptation of Large Models
- Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression
- Revisiting the Reliability of Psychological Scales on Large Language Models
- Who is ChatGPT? Benchmarking LLMs' Psychological Portrayal Using PsychoBench
- SyPS: Measuring Sycophancy Prompt Sensitivity in Large Language Models
- Is ChatGPT a Good Personality Recognizer? A Preliminary Study
- Evaluating and Inducing Personality in Pre-trained Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering