Latent Confidence Alignment for LLM Self-Assessment

πŸ“… 2026-06-20
πŸ“ˆ Citations: 0
✨ Influential: 0
πŸ“„ PDF
πŸ€– AI Summary
This work addresses the limitation in current large language models (LLMs) where confidence calibration often disregards item difficulty, making it difficult to discern whether high confidence stems from genuine self-assessment or artifacts of the generation process. To tackle this, the study introduces a novel framework that incorporates item difficulty into self-evaluation by leveraging the Rasch model to construct a latent ability space. Adopting a metacognitive perspective, it proposes the Latent Confidence Alignment Error (LCAE) metric to quantify the consistency between a model’s self-reported confidence and the error probability implied by its latent ability and item difficulty. Evaluation across 20 LLMs on a medical dataset demonstrates that this approach significantly enhances self-evaluation reliability without compromising task performance, while also uncovering a link between model reliability and reasoning cost.
πŸ“ Abstract
Confidence calibration in large language models (LLMs) is commonly evaluated by comparing predicted confidence with observed accuracy. However, such approaches do not model item difficulty, making it difficult to interpret discrepancies and to determine whether model confidence reflects genuine self-assessment or is merely a byproduct of the response generation process. To address this, we adopt a Rasch model-based latent ability framework and a metacognitive perspective, and propose Latent Confidence Alignment Error (LCAE) to measure the consistency between model self-assessment and the latent error probability implied by model ability and item difficulty. We further incorporate item difficulty as an external signal with a reasoning mechanism. Experiments on a medical-domain dataset with 20 models show that the proposed approach improves self-assessment quality without affecting model ability, and reveals an association between reliability and inference cost.
Problem

Research questions and friction points this paper is trying to address.

confidence calibration
self-assessment
item difficulty
large language models
latent ability
Innovation

Methods, ideas, or system contributions that make the work stand out.

Latent Confidence Alignment
Rasch model
self-assessment
item difficulty
confidence calibration
πŸ”Ž Similar Papers
No similar papers found.
πŸ’Ό Related Jobs
No related jobs found.
T
Ting-Yu Chen
Department of Information Management, National Sun Yat-Sen University, Kaohsiung, Taiwan
Tingting Yu
Tingting Yu
Associate Professor, University of Connecticut
Software EngineeringSoftware Testing
P
Pei-Cing Huang
Department of Information Management, National Sun Yat-Sen University, Kaohsiung, Taiwan
Chan Hsu
Chan Hsu
Ph.D. Student at National Sun Yat-sen University
Machine LearningInterpretabilityCausality
M
Ming-Yen Lin
Kaohsiung Medical University Hospital, Kaohsiung Medical University, Kaohsiung, Taiwan
Yihuang Kang
Yihuang Kang
National Sun Yat-sen University
Statistical Machine LearningHealth Services ResearchHealth Informatics