Pinocchio: Fast Uncertainty Estimates for Black-Box Language Models
cs.AI
Submitted: 2026-09-21
Updated: 2026-09-22
Code: https://github.com/khayes95/pinocchio
License: http://creativecommons.org/licenses/by/4.0/
The gist: In high-stakes decision-making applications of large language models (LLMs), practitioners require not only accurate LLMs but also uncertainty estimates for their predictions.
Terminology
Abstract
In high-stakes decision-making applications of large language models (LLMs), practitioners require not only accurate LLMs but also uncertainty estimates for their predictions. Existing approaches to uncertainty estimation for LLMs require access to log-probabilities output by the model or require fine-tuning access. However, many industrial LLM products use closed-source API models, and many such API models like GPT do not return log-probabilities and may not allow fine-tuning. We introduce Pinocchio, an external calibrator that estimates the correctness of responses from black-box API models. Trained jointly on responses from seven LLMs, it achieves 0.862 AUROC predicting the correctness of held-out responses from those same models, and shows zero-shot transfer to thirteen unseen models across eight organizations. Our model needs only a single forward pass to generate an uncertainty estimate and requires no access to the target model's logits, weights, or internal states. A lightweight text only 0.8B checkpoint matches our largest model's AUROC. We release code for adding uncertainty estimation to existing repos in only two additional lines of code.
Sources
- Qwen3-VL Technical Report
- Are We on the Right Way for Evaluating Large Vision-Language Models?
- Training Verifiers to Solve Math Word Problems
- Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
- Language Models (Mostly) Know What They Know
- BIG-Bench Extra Hard
- Are large language models superhuman chemists?
- Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
- Humanity's Last Exam
- MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
- Uncertainty-o: One Model-agnostic Framework for Unveiling Uncertainty in Large Multimodal Models
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection