EventDance: Unsupervised Source-Free Cross-Modal Adaptation for Event-Based Object Recognition
Xu Zheng, Lin Wang
Abstract
In this paper, we make the first attempt at achieving the cross-modal (i.e., image-to-events) adaptation for event-based object recognition without accessing any labeled source image data owning to privacy and commercial issues. Tackling this novel problem is non-trivial due to the novelty of event cameras and the distinct modality gap between images and events. In particular, as only the source model is available, a hurdle is how to extract the knowledge from the source model by only using the unlabeled target event data while achieving knowledge transfer. To this end, we propose a novel framework, dubbed Event-Dance for this unsupervised source-free cross-modal adaptation problem. Importantly, inspired by event-to-video reconstruction methods, we propose a reconstruction-based modality bridging (RMB) module, which reconstructs intensity frames from events in a self-supervised manner. This makes it possible to build up the surrogate images to extract the knowledge (i.e., labels) from the source model. We then propose a multi-representation knowledge adaptation (MKA) module that transfers the knowledge to target models learning events with multiple representation types for fully exploring the spatiotemporal information of events. The two modules connecting the source and target models are mutually updated so as to achieve the best performance. Experiments on three benchmark datasets with two adaption settings show that EventDance is on par with prior methods utilizing the source data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 20ca534a-7a1c-48d4-bad0-bbb1f9411321Cited by top-tier papers8
- Efficient Event Camera Data Pretraining with Adaptive Prompt FusionQuanmin Liang, Qiang Li, Shuai Liu, Xinzi Cao et al.ICCV 2025 · 6 citations
- Depth Any Event Stream: Enhancing Event-based Monocular Depth Estimation via Dense-to-Sparse DistillationJinjing Zhu, Tianbo Pan, Zidong Cao, Yexin Liu et al.ICCV 2025 · 3 citations
- From Sharp to Blur: Unsupervised Domain Adaptation for 2D Human Pose Estimation Under Extreme Motion Blur Using Event CamerasYoungho Kim, Hoonhee Cho, Kuk-Jin YoonICCV 2025 · 2 citations
- Multimodal Decomposed Distillation with Instance Alignment and Uncertainty Compensation for Thermal Object DetectionYanfeng Liu, Lefei ZhangACM MM 2025 · 2 citations
- Reducing Unimodal Bias in Multi-Modal Semantic Segmentation With Multi-Scale Functional Entropy RegularizationXu Zheng, Yuanhuiyi Lyu, Lutao Jiang, Danda Pani Paudel et al.ICCV 2025 · 2 citations
Builds on24
- Do We Really Need to Access the Source Data? Source Hypothesis Transfer for Unsupervised Domain AdaptationJian Liang, Dapeng Hu, Jiashi FengICML 2020 · 1,624 citations
- End-to-End Learning of Representations for Asynchronous Event-Based DataDaniel Gehrig, Antonio Loquercio, Konstantinos G. Derpanis, Davide ScaramuzzaICCV 2019 · 427 citations
- Category Contrast for Unsupervised Domain Adaptation in Visual TasksJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian Lu et al.CVPR 2022 · 143 citations
- Source-Free Domain Adaptation via Distribution EstimationNing Ding, Yixing Xu, Yehui Tang, Chao Xu et al.CVPR 2022 · 134 citations
- The Norm Must Go On: Dynamic Unsupervised Domain Adaptation by NormalizationMuhammad Jehanzeb Mirza, Jakub Micorek, Horst Possegger, Horst BischofCVPR 2022 · 119 citations
Related papers
- CMDA: Cross-Modality Domain Adaptation for Nighttime Semantic SegmentationRuihao Xia, Chaoqiang Zhao, Meng Zheng, Ziyan Wu et al.ICCV 2023 · 54 citations
- EvDistill: Asynchronous Events To End-Task Learning via Bidirectional Reconstruction-Guided Cross-Modal Knowledge DistillationLin Wang, Yujeong Chae, Sung-Hoon Yoon, Tae-Kyun Kim et al.CVPR 2021
- Bridge Then Begin Anew: Generating Target-Relevant Intermediate Model for Source-Free Visual Emotion AdaptationJiankun Zhu, Sicheng Zhao, Jing Jiang, Wenbo Tang et al.AAAI 2025 · 1 citation
- Learning Adaptive Dense Event Stereo from the Image DomainHoonhee Cho, Jegyeong Cho, Kuk-Jin YoonCVPR 2023
- CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training FrameworkWentao Wu, Xiao Wang, Chenglong Li, Bo Jiang et al.ACM MM 2025 · 2 citations
