Continual Momentum Filtering on Parameter Space for Online Test-time Adaptation
Jae-Hong Lee, Joon-Hyuk Chang
Abstract
Deep neural networks (DNNs) have revolutionized tasks such as image classification and speech recognition but often falter when training and test data diverge in distribution. External factors, from weather effects on images to varied speech environments, can cause this discrepancy, compromising DNN performance. Online test-time adaptation (OTTA) methods present a promising solution, recalibrating models in real-time during the test stage without requiring historical data. However, the OTTA paradigm is imperfect, often falling prey to issues such as catastrophic forgetting due to its reliance on noisy, self-trained predictions. Although some contemporary strategies mitigate this by tying adaptations to the static source model, this restricts model flexibility. This paper introduces a continual momentum filtering (CMF) framework, leveraging the Kalman filter (KF) to strike a balance between model adaptability and information retention. The CMF intertwines optimization via stochastic gradient descent with a KF-based inference process. This methodology not only aids in averting catastrophic forgetting but also provides high adaptability to shifting data distributions. We validate our framework on various OTTA scenarios and real-world situations regarding covariate and label shifts, and the CMF consistently shows superior performance compared to state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 260178ae-89c7-40bc-9cd4-49d7a3583eecCited by top-tier papers8
- Buffer layers for Test-Time AdaptationHyeongyu Kim, Geonhui Han, Dosik HwangNeurIPS 2025 · 6 citations
- When and Where to Reset Matters for Long-Term Test-Time AdaptationTaejun Lim, Joong-Won Hwang, Kibok LeeICLR 2026 · 3 citations
- Hybrid-Tta: Continual Test-Time Adaptation Via Dynamic Domain Shift DetectionHyewon Park, Hyejin Park, Jueun Ko, Dongbo MinICCV 2025 · 2 citations
- Tempora: Characterising the Time-Contingent Utility of Online Test-Time AdaptationSudarshan Sreeram, Young D. Kwon, Cecilia MascoloICML 2026 · 1 citation
- AcTTA: Rethinking Test-Time Adaptation via Dynamic ActivationHyeongyu Kim, Geonhui Han, Dosik HwangCVPR 2026
Builds on23
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 9,451 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
Related papers
- Continual Test-Time Domain AdaptationQin Wang, Olga Fink, Luc Van Gool, Dengxin DaiCVPR 2022 · 383 citations
- Stationary Latent Weight Inference for Unreliable Observations from Online Test-Time AdaptationJae-Hong Lee, Joon-Hyuk ChangICML 2024 · 6 citations
- MINGLE: Mixture of Null-Space Gated Low-Rank Experts for Test-Time Continual Model MergingZihuan Qiu, Yi Xu, Chiyuan He, Fanman Meng et al.NeurIPS 2025 · 16 citations
- Overcoming Recency Bias of Normalization Statistics in Continual Learning: Balance and AdaptationYilin Lyu, Liyuan Wang, Xingxing Zhang, Zicheng Sun et al.NeurIPS 2023 · 17 citations
- Kalman Filter for Online Classification of Non-Stationary DataMichalis K. Titsias, Alexandre Galashov, Amal Rannen-Triki, Razvan Pascanu et al.ICLR 2024 · 14 citations
