IT3: Idempotent Test-Time Training
Nikita Durasov, Assaf Shocher, Doruk Öner, Gal Chechik, Alexei A. Efros, Pascal Fua
摘要
Deep learning models often struggle when deployed in real-world settings due to distribution shifts between training and test data. While existing approaches like domain adaptation and test-time training (TTT) offer partial solutions, they typically require additional data or domainspecific auxiliary tasks. We present Idempotent Test-Time Training (IT 3 ), a novel approach that enables on-the-fly adaptation to distribution shifts using only the current test instance, without any auxiliary task design. Our key insight is that enforcing idempotence-where repeated applications of a function yield the same result-can effectively replace domain-specific auxiliary tasks used in previous TTT methods. We theoretically connect idempotence to prediction confidence and demonstrate that minimizing the distance between successive applications of our model during inference leads to improved out-of-distribution performance. Extensive experiments across diverse domains (including image classification, aerodynamics prediction, and aerial segmentation) and architectures (MLPs, CNNs, GNNs) show that IT 3 consistently outperforms existing approaches while being simpler and more widely applicable. Our results suggest that idempotence provides a universal principle for test-time adaptation that generalizes across domains and architectures. poster / code / video / web
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Specialization after Generalization: Towards Understanding Test-Time Training in Foundation ModelsJonas Hübotter, Patrik Wolf, Aleksandr Shevchenko, Dennis Jüni 等ICLR 2026 · 被引用 6 次
- Multivariate Time Series Anomaly Detection with Idempotent ReconstructionXin Sun, Heng Zhou, Chao LiNeurIPS 2025 · 被引用 5 次
- IDER: IDempotent Experience Replay for Reliable Continual LearningZhanwang Liu, Yuting Li, Haoyuan Gao, Yexin Li 等ICLR 2026 · 被引用 5 次
- Who Said Neural Networks Aren't Linear?Nimrod Berman, Assaf Hallak, Assaf ShocherICML 2026 · 被引用 3 次
- A Provable Energy-Guided Test-Time Defense Boosting Adversarial Robustness of Large Vision-Language ModelsMujtaba Hussain Mirza, Antonio D’Orazio, Odelia Melamed, Iacopo MasiCVPR 2026 · 被引用 2 次
它引用的顶会 Paper14
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath 等ICCV 2021 · 被引用 2,294 次
- Tent: Fully Test-Time Adaptation by Entropy MinimizationDequan Wang, Evan Shelhamer, Shaoteng Liu, Bruno A. Olshausen 等ICLR 2021 · 被引用 1,731 次
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller 等ICML 2020 · 被引用 1,220 次
- Efficient Test-Time Model Adaptation without ForgettingShuaicheng Niu, Jiaxiang Wu, Yifan Zhang, Yaofo Chen 等ICML 2022 · 被引用 579 次
- TTT++: When Does Self-Supervised Test-Time Training Fail or Thrive?Yuejiang Liu, Parth Kothari, Bastien van Delft, Baptiste Bellot-Gurlet 等NeurIPS 2021 · 被引用 469 次
相关 Paper
- Improved Test-Time Adaptation for Domain GeneralizationLiang Chen, Yong Zhang, Yibing Song, Ying Shan 等CVPR 2023
- Synchronizing Task Behavior: Aligning Multiple Tasks During Test-Time TrainingWooseong Jeong, Jegyeong Cho, Youngho Yoon, Kuk-Jin YoonICCV 2025
- ClusT3: Information Invariant Test-Time TrainingGustavo Adolfo Vargas Hakim, David Osowiechi, Mehrdad Noori, Milad Cheraghalikhani 等ICCV 2023 · 被引用 25 次
- MATE: Masked Autoencoders are Online 3D Test-Time LearnersMuhammad Jehanzeb Mirza, Inkyu Shin, Wei Lin, Andreas Schriebl 等ICCV 2023 · 被引用 24 次
- Test-Time Classifier Adjustment Module for Model-Agnostic Domain GeneralizationYusuke Iwasawa, Yutaka MatsuoNeurIPS 2021 · 被引用 456 次
