Lune

ICML2026顶会

Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers

Jingkai Huang, Will Ma, Zhengyuan Zhou

2026年份
3被引次数

摘要

A simple strategy for improving LLM accuracy, especially in math and reasoning problems, is to sample multiple responses and submit the answer most consistently reached. In this paper we leverage Bayesian prior information to save on sampling costs, stopping once sufficient consistency is reached. Although the exact posterior is computationally intractable, we further introduce an efficient ``LL-aggregated'' stopping policy that tracks only the L−1L-1 most frequent answer counts. Theoretically, we prove that L=3L=3 is all you need: this coarse approximation is sufficient to achieve asymptotic optimality, and strictly dominates prior-free baselines, while having a fast posterior computation. Empirically, this identifies the most consistent (i.e., mode) LLM answer and achieves similar answer accuracy using fewer samples.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext a43f3b3c-e1e6-492b-89e8-5b1cd56367ec

它引用的顶会 Paper10

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖