Lune

ICML2026顶会

Calibrated Knowledge Aggregation in Bayesian Mixture-of-Experts for Continual VQA

Mahsa Mozaffari, Hitesh Sapkota, Yu Kong, Xumin Liu, Qi Yu

出版方
2026年份

摘要

Continual learning for visual question answering (VQA) is typically implemented by training one expert per task and routing each query using task-ID supervision. Yet continual VQA tasks overlap substantially: on the VQA-v2 task stream, a non-native expert outperforms the task’s own expert on 49.9%49.9\% of queries, so hard routing both wastes transferable knowledge and can be confidently wrong when mismatched. We propose a calibrated Bayesian mixture-of-experts that trains parameter-efficient per-task adapters, learns routing by directly maximizing expected VQA utility, and marginalizes expert identity at inference via Bayesian aggregation in a unified answer space; an entropy penalty prevents the utility objective from collapsing to one-hot routing, enabling evidence pooling across plausible experts. We reach 64.1664.16 accuracy with 0.630.63 forgetting on VQA-v2 CL-LS (+5.74%+5.74\% accuracy, −2.99-2.99 forgetting vs. the strongest prior method), 78.8178.81 with 0.400.40 forgetting on TDIUC CL-LS (+3.10+3.10, −1.74-1.74), and 83.4183.41 with 3.213.21 forgetting on TDIUC CL-VS (+1.58+1.58, −0.82-0.82). Calibration also improves on VQA-v2, reducing ECE from 0.150.15 to 0.070.07.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper19

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖