MECTA: Memory-Economic Continual Test-Time Model Adaptation
Junyuan Hong, Lingjuan Lyu, Jiayu Zhou, Michael Spranger
摘要
Continual Test-time Adaptation (CTA) is a promising art to secure accuracy gains in continually-changing environments. The state-of-the-art adaptations improve out-of-distribution model accuracy via computation-efficient online test-time gradient descents but meanwhile cost about times of memory versus the inference, even if only a small portion of parameters are updated. Such high memory consumption of CTA substantially impedes wide applications of advanced CTA on memory-constrained devices. In this paper, we provide a novel solution, dubbed MECTA, to drastically improve the memory efficiency of gradient-based CTA. Our profiling shows that the major memory overhead comes from the intermediate cache for back-propagation, which scales by the batch size, channel, and layer number. Therefore, we propose to reduce batch sizes, adopt an adaptive normalization layer to maintain stable and accurate predictions, and stop the back-propagation caching heuristically. On the other hand, we prune the networks to reduce the computation and memory overheads in optimization and recover the parameters afterward to avoid forgetting. The proposed MECTA is efficient and can be seamlessly plugged into state-of-the-art CTA algorithms at negligible overhead on computation and memory. On three datasets, CIFAR10, CIFAR100, and ImageNet, MECTA improves the accuracy by at least 6% with constrained memory and significantly reduces the memory costs of ResNet50 on ImageNet by at least 70% with comparable accuracy. Our codes can be accessed at https://github.com/SonyAI/MECTA.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper25
- SoTTA: Robust Test-Time Adaptation on Noisy Data StreamsTaesik Gong, Yewon Kim, Taeckyung Lee, Sorn Chottananurak 等NeurIPS 2023 · 被引用 89 次
- Test-Time Model Adaptation with Only Forward PassesShuaicheng Niu, Chunyan Miao, Guohao Chen, Pengcheng Wu 等ICML 2024 · 被引用 77 次
- Towards Stable Test-time Adaptation in Dynamic Wild WorldShuaicheng Niu, Jiaxiang Wu, Yifan Zhang, Zhiquan Wen 等ICLR 2023 · 被引用 62 次
- TinyTTA: Efficient Test-time Adaptation via Early-exit Ensembles on Edge DevicesHong Jia, Young D. Kwon, Alessio Orsino, Ting Dang 等NeurIPS 2024 · 被引用 25 次
- L-TTA: Lightweight Test-Time Adaptation Using a Versatile Stem LayerJin Shin, Hyun KimNeurIPS 2024 · 被引用 23 次
相关 Paper
- EcoTTA: Memory-Efficient Continual Test-Time Adaptation via Self-Distilled RegularizationJunha Song, Jungsoo Lee, In So Kweon, Sungha ChoiCVPR 2023
- SNAP: Low-Latency Test-Time Adaptation with Sparse UpdatesHyeongheon Cha, Dong Min Kim, Hye Won Chung, Taesik Gong 等NeurIPS 2025 · 被引用 4 次
- FOZO: Forward-Only Zeroth-Order Prompt Optimization for Test-Time AdaptationXingyu Wang, Tao WangCVPR 2026 · 被引用 2 次
- EBaR: Efficient Buffer and Resetting for Single-Sample Continual Test-Time AdaptationTianyi Ma, Maoying QiaoACM MM 2025
- Continual Test-Time Domain AdaptationQin Wang, Olga Fink, Luc Van Gool, Dengxin DaiCVPR 2022 · 被引用 383 次
