Shared Doubt: Zero-Shot Cross-Lingual Confidence Estimation for Language Models
cs.CL, cs.AI, cs.LG
Submitted: 2026-05-29
Updated: 2026-09-27
Code: https://github.com/AthinaKyriakou/shared-doubt
Terminology
Sources
- Uncertainty in Natural Language Generation: From Theory to Applications
- Fine-Grained Uncertainty Quantification for Long-Form Language Model Outputs: A Comparative Study
- Discovering Latent Knowledge in Language Models Without Supervision
- Answer Matching Outperforms Multiple Choice for Language Model Evaluation
- Do LLMs Know about Hallucination? An Empirical Investigation of LLM's Hidden States
- IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
- A Primer on the Inner Workings of Transformer-based Language Models
- The Llama 3 Herd of Models
- A Survey of Uncertainty Estimation in LLMs: Theory Meets Practice
- On the Origins of Linear Representations in Large Language Models
- Language Models (Mostly) Know What They Know
- Exploring Multilingual Probing in Large Language Models: A Cross-Language Analysis
- Teaching Models to Express Their Uncertainty in Words
- Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
- LitCab: Lightweight Language Model Calibration over Short- and Long-form Responses
- Confidence over Time: Confidence Calibration with Temporal Logic for Large Language Model Reasoning
- The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets
- The Linear Representation Hypothesis and the Geometry of Large Language Models
- Qwen3 Technical Report
- Do Multilingual LLMs Think In English?
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering