Calibrated Language Models and How to Find Them with Label Smoothing
Jerry Huang, Peng Lu, Qiuhao Zeng
摘要
Recent advances in natural language processing have enabled the fine-tuning of large language models (LLMs) into powerful interactive agents with improved instruction-following ability. However, this can impact confidence calibration for reliable model output, which has not been researched in full. In this work, we examine various open-sourced LLMs, where we identify significant calibration degradation after instruction tuning. Seeking a practical solution, we look towards label smoothing, which has been shown as an effective method to regularize for overconfident predictions but has yet to be widely adopted in the supervised fine-tuning (SFT) of LLMs. We provide insight into why label smoothing can maintain calibration throughout the SFT process, but identify settings remain where the effectiveness of smoothing is severely diminished. We posit the cause to stem from the ability to become overconfident, which has a direct relationship with the hidden and vocabulary size of models, which we justify theoretically and experimentally. Finally, we address an outstanding issue regarding the memory footprint of the cross-entropy loss computation with label smoothing, designing a customized kernel to dramatically reduce memory consumption without sacrificing speed or performance in comparison to existing solutions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Mamba Modulation: On the Length Generalization of Mamba ModelsPeng Lu, Jerry Huang, Qiuhao Zeng, Xinyu Wang 等NeurIPS 2025 · 被引用 2 次
- Confidence is Not Universal: Task-Dependent Calibration and Emergent Behavior in LLMsChaeyun Jang, Moonseok Choi, Yegon Kim, Seungyoo Lee 等ICML 2026
- Attention with Routed-Memory for Learnable Sparse ControlQIUHAO Zeng, Jerry Huang, Peng Lu, Ruiyi Fang 等ICML 2026
它引用的顶会 Paper20
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou 等ICLR 2021 · 被引用 7,905 次
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra 等NeurIPS 2022 · 被引用 5,493 次
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu 等ICLR 2022 · 被引用 4,966 次
- FlashAttention-2: Faster Attention with Better Parallelism and Work PartitioningTri DaoICLR 2024 · 被引用 2,600 次
相关 Paper
- A Close Look into the Calibration of Pre-trained Language ModelsYangyi Chen, Lifan Yuan, Ganqu Cui, Zhiyuan Liu 等ACL 2023 · 被引用 12 次
- SFTMix: Elevating Language Model Instruction Tuning with Mixup RecipeYuxin Xiao, Shujian Zhang, Marzyeh Ghassemi, Wenxuan ZhouACL 2026 · 被引用 3 次
- Adaptive Label Smoothing with Self-Knowledge in Natural Language GenerationDongkyu Lee, Ka Chun Cheung, Nevin L. ZhangEMNLP 2022 · 被引用 5 次
- AgentRefine: Enhancing Agent Generalization through Refinement TuningDayuan Fu, Keqing He, Yejie Wang, Wentao Hong 等ICLR 2025
- Enhancing Language Model Alignment: A Confidence-Based Approach to Label SmoothingBaihe Huang, Hiteshi Sharma, Yi MaoEMNLP 2024 · 被引用 1 次
