Robust Object Detection via Instance-Level Temporal Cycle Confusion
Xin Wang, Thomas E. Huang, Benlin Liu, Fisher Yu, Xiaolong Wang, Joseph E. Gonzalez, Trevor Darrell
Abstract
Building reliable object detectors that are robust to domain shifts, such as various changes in context, viewpoint, and object appearances, is critical for real-world applications. In this work, we study the effectiveness of auxiliary self-supervised tasks to improve the out-of-distribution generalization of object detectors. Inspired by the principle of maximum entropy, we introduce a novel self-supervised task, instance-level temporal cycle confusion (CycConf), which operates on the region features of the object detectors. For each object, the task is to find the most different object proposals in the adjacent frame in a video and then cycle back to itself for self-supervision. CycConf encourages the object detector to explore invariant structures across instances under various motions, which leads to improved model robustness in unseen domains at test time. We observe consistent out-of-domain performance improvements when training object detectors in tandem with self-supervised tasks on various do-main adaptation benchmarks with static images (Cityscapes, Foggy Cityscapes, Sim10K) and large-scale video datasets (BDD100K and Waymo open data)1.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 467c4495-5f9c-443a-a90a-3aa02bb6c283Cited by top-tier papers11
- VOS: Learning What You Don't Know by Virtual Outlier SynthesisXuefeng Du, Zhaoning Wang, Mu Cai, Yixuan LiICLR 2022 · 417 citations
- Unknown-Aware Object Detection: Learning What You Don't Know from Videos in the WildXuefeng Du, Xin Wang, Gabriel Gozum, Yixuan LiCVPR 2022 · 70 citations
- Pasta: Proportional Amplitude Spectrum Training Augmentation for Syn-to-Real Domain GeneralizationPrithvijit Chattopadhyay, Kartik Sarangmath, Vivek Vijaykumar, Judy HoffmanICCV 2023 · 57 citations
- H2FA R-CNN: Holistic and Hierarchical Feature Alignment for Cross-domain Weakly Supervised Object DetectionYunqiu Xu, Yifan Sun, Zongxin Yang, Jiaxu Miao et al.CVPR 2022 · 40 citations
- PhysAug: A Physical-guided and Frequency-based Data Augmentation for Single-Domain Generalized Object DetectionXiaoran Xu, Jiangang Yang, Wenhui Shi, Siyuan Ding et al.AAAI 2025 · 15 citations
Builds on8
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- AugMix: A Simple Data Processing Method to Improve Robustness and UncertaintyDan Hendrycks, Norman Mu, Ekin Dogus Cubuk, Barret Zoph et al.ICLR 2020 · 1,572 citations
- Multi-Adversarial Faster-RCNN for Unrestricted Object DetectionZhenwei He, Lei ZhangICCV 2019 · 352 citations
- Maximum Entropy RL (Provably) Solves Some Robust RL ProblemsBenjamin Eysenbach, Sergey LevineICLR 2022 · 244 citations
- Cross-Domain Detection via Graph-Induced Prototype AlignmentMinghao Xu, Hang Wang, Bingbing Ni, Qi Tian et al.CVPR 2020
Related papers
- Single-Domain Generalized Object Detection in Urban Scene via Cyclic-Disentangled Self-DistillationAming Wu, Cheng DengCVPR 2022 · 110 citations
- Object Detection with Self-Supervised Scene AdaptationZekun Zhang, Minh HoaiCVPR 2023
- SSAL: Synergizing between Self-Training and Adversarial Learning for Domain Adaptive Object DetectionMuhammad Akhtar Munir, Muhammad Haris Khan, M. Saquib Sarfraz, Mohsen AliNeurIPS 2021 · 53 citations
- Towards Building Self-Aware Object Detectors via Reliable Uncertainty Quantification and CalibrationKemal Oksuz, Tom Joy, Puneet K. DokaniaCVPR 2023
- Self-Supervised Multi-Object Tracking with Cross-input ConsistencyFavyen Bastani, Songtao He, Samuel MaddenNeurIPS 2021 · 39 citations
