Large Language Models Exhibit Human-Like Bayesian Hypocrisy
cs.HC, cs.CL, cs.CY
Submitted: 2026-08-03
Updated: 2026-08-03
Terminology
Sources
- Constitutional AI: Harmlessness from AI Feedback
- ARC-AGI-2: A New Challenge for Frontier AI Reasoning Systems
- Scaling Laws for Neural Language Models
- Training language models to follow instructions with human feedback
- The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
- Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks
- Verbalized Sampling: How to Mitigate Mode Collapse and Unlock LLM Diversity
Related papers
- EduGage: A Multimodal Dataset and Benchmark for Sensor-Based Momentary Assessment of Engagement in Self-Guided Video Learning
- EvoDesign: Agentic Editable Diagram Creation via Design Expertise Evolution
- HAGI++: Head-Assisted Gaze Imputation and Generation
- Linking Behaviour and Perception to Evaluate Meaningful Human Control over Partially Automated Driving
- Review of Explainable Decision Support and Adaptive Human-Machine Interfaces for Automation Transparency in Maritime Autonomous Surface Ships
- Towards Cognitive Process-Aware Proactive Writing Support