Towards Reliable Medical LLMs: Benchmarking and Enhancing Confidence Estimation of Large Language Models in Medical Consultation

arXiv:2601.15645 · cs.CL · Submitted 2026-01-22 · Read on arXiv

cs.CL

Submitted: 2026-01-22

Updated: 2026-09-10

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers