LLMs1 min read
Analysis of calibration issues in language model verdicts and internal knowledge
A 0.6B language model's behavior was examined, revealing calibration failures where internal verdicts are misaligned with output logits, affecting accuracy and interpretability.
From arXiv cs.CL