Rare Event Analysis of Large Language Models
Jake McAllister Dorman, Edward Gillman, Dominic C Rose, Jamie Mair, Juan Garrahan
摘要
Being probabilistic models, during inference large language models (LLMs) display rare events : behaviour that is far from typical but highly significant. By definition all rare events are hard to see, but the enormous scale of LLM usage means that events completely unobserved during development are likely to become prominent in deployment. Here we present an end-to-end framework for the systematic analysis of rare events in LLMs. We provide a practical implementation spanning theory, efficient generation strategies, probability estimation and error analysis, which we illustrate with concrete examples. We outline extensions and applications to other models and contexts, highlighting the generality of the concepts and techniques presented here.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper10
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes 等ICLR 2020 · 被引用 4,112 次
- HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust RefusalMantas Mazeika, Long Phan, Xuwang Yin, Andy Zou 等ICML 2024 · 被引用 1,031 次
- Reasoning with Sampling: Your Base Model is Smarter Than You ThinkAayush Karan, Yilun DuICLR 2026 · 被引用 87 次
- A Thorough Examination of Decoding Methods in the Era of LLMsChufan Shi, Haoran Yang, Deng Cai, Zhisong Zhang 等EMNLP 2024 · 被引用 26 次
相关 Paper
- Evaluating Distributional Distortion in Neural Language ModelingBenjamin LeBrun, Alessandro Sordoni, Timothy J. O'DonnellICLR 2022 · 被引用 26 次
- Mining Long Tail Bugs: Identifying Rare and Overlooked Issues in CodeWentao Liang, Yanjun Wu, Xiang Ling, Tianyue Luo 等FSE 2026
- Are Language Models Any Good at Density Modeling?Sriram Ranga, Sai Shashank Bedampeta, Rui Mao, Anupam ChattopadhyayAAAI 2026
- FT2: First-Token-Inspired Online Fault Tolerance on Critical Layers for Generative Large Language ModelsYu Sun, Zhu Zhu, Cherish Mulpuru, Roberto Gioiosa 等HPDC 2025 · 被引用 7 次
- What is a protest anyway? Codebook conceptualization is still a first-order concern in LLM-era classificationAndrew Halterman, Katherine A. KeithACL 2026 · 被引用 3 次
