Min-k Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics
Yuanhao Ding, Meimingwei Li, Esteban Garces Arias, Matthias Aßenmacher, Christian Heumann, Chongsheng Zhang
Abstract
The quality of text generated by large language models depends critically on the decoding sampling strategy. While mainstream methods such as Top-, Top-, and Min- achieve a balance between diversity and accuracy through probability-space truncation, they share an inherent limitation: extreme sensitivity to the temperature parameter. Recent logit-space approaches like Top- achieve temperature invariance but rely on global statistics that are susceptible to long-tail noise, failing to capture fine-grained confidence structures among top candidates. We propose Min- Sampling, a novel dynamic truncation strategy that analyzes the local shape of the sorted logit distribution to identify"semantic cliffs": sharp transitions from high-confidence core tokens to uncertain long-tail tokens. By computing a position-weighted relative decay rate, Min- dynamically determines truncation boundaries at each generation step. We formally prove that Min- achieves strict temperature invariance and empirically demonstrate its low sensitivity to hyperparameter choices. Experiments on multiple reasoning benchmarks, creative writing tasks, and human evaluation show that Min- consistently improves text quality, maintaining robust performance even under extreme temperature settings where probability-based methods collapse. We make our code, models, and analysis tools publicly available.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 81a5a7df-d4e8-4d47-8a7c-6e28302eba78Cited by top-tier papers2
- Digitizing Nepal's Written Heritage: A Comprehensive HTR Pipeline for Old Nepali ManuscriptsAnjali Sarawgi, Esteban Garces Arias, Christof ZotterACL 2026 · 2 citations
- Beyond Temperature: Hyperfitting as a Late-Stage Geometric ExpansionMeimingwei Li, Yuanhao Ding, Esteban Garces Arias, Christian HeumannICML 2026
Builds on8
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes et al.ICLR 2020 · 4,112 citations
- Let's Verify Step by StepHunter Lightman, Vineet Kosaraju, Yuri Burda, Harrison Edwards et al.ICLR 2024 · 3,045 citations
- Improving Open-Ended Text Generation via Adaptive DecodingWenhong Zhu, Hongkun Hao, Zhiwei He, Yiming Ai et al.ICML 2024 · 21 citations
- Mirostat: a Neural Text decoding Algorithm that directly controls perplexitySourya Basu, Govardana Sachitanandam Ramachandran, Nitish Shirish Keskar, Lav R. VarshneyICLR 2021 · 13 citations
- Top-nσ: Eliminating Noise in Logit Space for Robust Token Sampling of LLMChenxia Tang, Jianchun Liu, Hongli Xu, Liusheng HuangACL 2025 · 9 citations
Related papers
- Turning Up the Heat: Min-p Sampling for Creative and Coherent LLM OutputsNguyen Nhat Minh, Andrew Baker, Clement Neo, Allen G. Roush et al.ICLR 2025
- p-less Sampling: A Robust Hyperparameter-Free Approach for LLM DecodingRunyan Tan, Shuang Wu, Phillip HowardICLR 2026 · 2 citations
- Top-H Decoding: Adapting the Creativity and Coherence with Bounded Entropy in Text GenerationErfan Baghaei Potraghloo, Seyedarmin Azizi, Souvik Kundu, Massoud PedramNeurIPS 2025 · 13 citations
- Sample Smart, Not Hard: Correctness-First Decoding for Better Reasoning in LLMsXueyan Li, Guinan Su, Mrinmaya Sachan, Jonas GeipingICLR 2026 · 5 citations
- Closing the Curious Case of Neural Text DegenerationMatthew Finlayson, John Hewitt, Alexander Koller, Swabha Swayamdipta et al.ICLR 2024 · 31 citations
