Post Hoc Regression Refinement via Pairwise Rankings
Kevin Tirta Wijaya, Michael Sun, Minghao Guo, Hans-Peter Seidel, Wojciech Matusik, Vahid Babaei
摘要
Accurate prediction of continuous properties is essential to many scientific and engineering tasks. Although deep-learning regressors excel with abundant labels, their accuracy deteriorates in data-scarce regimes. We introduce RankRefine, a model-agnostic, plug-and-play post hoc method that refines regression with expert knowledge coming from pairwise rankings. Given a query item and a small reference set with known properties, RankRefine combines the base regressor's output with a rank-based estimate via inverse-variance weighting, requiring no retraining. In molecular property prediction task, RankRefine achieves up to 10% relative reduction in mean absolute error using only 20 pairwise comparisons obtained through a general-purpose large language model (LLM) with no finetuning. As rankings provided by human experts or general-purpose LLMs are sufficient for improving regression across diverse domains, RankRefine offers practicality and broad applicability, especially in low-data settings.
Lemma 3.2 (Variance of the rank-based estimate) Let ŷ * rank 0 minimize equation 2. Its variance is approximated by the inverse observed Fisher information [Ly et al., 2017],
Applying Theorem 3.1 with σ 2 rank from Lemma 3.2 yields ŷ * 0 .
We measure performance via the mean absolute error (MAE). For a folded Gaussian derived from a zero-mean Gaussian, MAE = 2/π σ, so
Corollary 3.2.1 Any informative ranker with finite variance (σ 2 rank < ∞) lowers the expected MAE after fusion.
More generally, letting α ∈ [0, 1] be the desired ratio between post-refinement and the original MAEs,
Regularization. If the ranker is biased yet over-confident (σ 2 rank ≪ σ 2 reg ), we temper its variance via σ 2 rank ← max σ 2 rank , c σ 2 reg , with user-chosen constant c > 0.
We evaluate RankRefine on synthetic and real-world tasks. After outlining datasets, metrics, and implementation details, we report results in synthetic settings, where ranking oracle is available, and practical settings across multiple domains.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- RankUp: Boosting Semi-Supervised Regression with an Auxiliary Ranking ClassifierPin-Yen Huang, Szu-Wei Fu, Yu TsaoNeurIPS 2024 · 被引用 13 次
- Representing Molecules as Random Walks Over Interpretable GrammarsMichael Sun, Minghao Guo, Weize Yuan, Veronika Thost 等ICML 2024 · 被引用 6 次
- Consolidating Ranking and Relevance Predictions of Large Language Models through Post-ProcessingLe Yan, Zhen Qin, Honglei Zhuang, Rolf Jagerman 等EMNLP 2024 · 被引用 6 次
- Foundation Molecular Grammar: Multi-Modal Foundation Models Induce Interpretable Molecular Graph LanguagesMichael Sun, Weize Yuan, Gang Liu, Wojciech Matusik 等ICML 2025
相关 Paper
- A Judge-Aware Ranking Framework for Evaluating Large Language Models without Ground TruthMingyuan Xu, Xinzi Tan, Jiawei Wu, Doudou ZhouICML 2026
- LLM Processes: Numerical Predictive Distributions Conditioned on Natural LanguageJames Requeima, John Bronskill, Dami Choi, Richard E. Turner 等NeurIPS 2024 · 被引用 72 次
- Curriculum Model Merging: Harmonizing Chemical LLMs for Enhanced Cross-Task GeneralizationBaoyi He, Luotian Yuan, Ying Wei, Fei WuNeurIPS 2025
- UDAPDR: Unsupervised Domain Adaptation via LLM Prompting and Distillation of RerankersJon Saad-Falcon, Omar Khattab, Keshav Santhanam, Radu Florian 等EMNLP 2023 · 被引用 23 次
- Open-World Planning via Lifted Regression with LLM-Inferred Affordances for Embodied AgentsXiaotian Liu, Ali Pesaranghader, Hanze Li, Punyaphat Sukcharoenchaikul 等ACL 2025 · 被引用 2 次
