EvDistill: Asynchronous Events To End-Task Learning via Bidirectional Reconstruction-Guided Cross-Modal Knowledge Distillation
Lin Wang, Yujeong Chae, Sung-Hoon Yoon, Tae-Kyun Kim, Kuk-Jin Yoon
Abstract
Event cameras sense per-pixel intensity changes and produce asynchronous event streams with high dynamic range and less motion blur, showing advantages over the conventional cameras. A hurdle of training event-based models is the lack of large qualitative labeled data. Prior works learning end-tasks mostly rely on labeled or pseudolabeled datasets obtained from the active pixel sensor (APS) frames; however, such datasets' quality is far from rivaling those based on the canonical images. In this paper, we propose a novel approach, called EvDistill, to learn a student network on the unlabeled and unpaired event data (target modality) via knowledge distillation (KD) from a teacher network trained with large-scale, labeled image data (source modality). To enable KD across the unpaired modalities, we first propose a bidirectional modality reconstruction (BMR) module to bridge both modalities and simultaneously exploit them to distill knowledge via the crafted pairs, causing no extra computation in the inference. The BMR is improved by the end-tasks and KD losses in an end-to-end manner. Second, we leverage the structural similarities of both modalities and adapt the knowledge by matching their distributions. Moreover, as most prior feature KD methods are uni-modality and less applicable to our problem, we propose an affinity graph KD loss to boost the distillation. Our extensive experiments on semantic segmentation and object recognition demonstrate that EvDistill achieves significantly better results than the prior works and KD with only events and APS frames.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f5a2440b-cf50-46b1-becf-a17d492a5be2Cited by top-tier papers26
- Cross-Domain Correlation Distillation for Unsupervised Domain Adaptation in Nighttime Semantic SegmentationHuan Gao, Jichang Guo, Guoli Wang, Qian ZhangCVPR 2022 · 82 citations
- A Voxel Graph CNN for Object Classification with Event CamerasYongjian Deng, Hao Chen, Hai Liu, Youfu LiCVPR 2022 · 55 citations
- CMDA: Cross-Modality Domain Adaptation for Nighttime Semantic SegmentationRuihao Xia, Chaoqiang Zhao, Meng Zheng, Ziyan Wu et al.ICCV 2023 · 54 citations
- Dual Transfer Learning for Event-based End-task Prediction via Pluggable Event to Image TranslationLin Wang, Yujeong Chae, Kuk-Jin YoonICCV 2021 · 46 citations
- Label-Free Event-based Object Recognition via Joint Learning with Image Reconstruction from EventsHoonhee Cho, Hyeonseong Kim, Yujeong Chae, Kuk-Jin YoonICCV 2023 · 38 citations
Builds on18
- A Comprehensive Overhaul of Feature DistillationByeongho Heo, Jeesoo Kim, Sangdoo Yun, Hyojin Park et al.ICCV 2019 · 727 citations
- End-to-End Learning of Representations for Asynchronous Event-Based DataDaniel Gehrig, Antonio Loquercio, Konstantinos G. Derpanis, Davide ScaramuzzaICCV 2019 · 427 citations
- Online Knowledge Distillation with Diverse PeersDefang Chen, Jian-Ping Mei, Can Wang, Yan Feng et al.AAAI 2020 · 354 citations
- Event-Based Motion Segmentation by Motion CompensationTimo Stoffregen, Guillermo Gallego, Tom Drummond, Lindsay Kleeman et al.ICCV 2019 · 164 citations
- Learning an Event Sequence Embedding for Dense Event-Based Deep StereoStepan Tulyakov, François Fleuret, Martin Kiefel, Peter V. Gehler et al.ICCV 2019 · 122 citations
Related papers
- Enhanced Event-Based Dense Stereo via Cross-Sensor Knowledge DistillationHaihao Zhang, Yunjian Zhang, Jianing Li, Lin Zhu et al.ICCV 2025 · 1 citation
- Event Stream-Based Visual Object Tracking: A High-Resolution Benchmark Dataset and A Novel BaselineXiao Wang, Shiao Wang, Chuanming Tang, Lin Zhu et al.CVPR 2024 · 48 citations
- Video to Events: Recycling Video Datasets for Event CamerasDaniel Gehrig, Mathias Gehrig, Javier Hidalgo-Carrió, Davide ScaramuzzaCVPR 2020
- EventDance: Unsupervised Source-Free Cross-Modal Adaptation for Event-Based Object RecognitionXu Zheng, Lin WangCVPR 2024 · 15 citations
- Distil-E2D: Distilling Image-to-Depth Priors for Event-Based Monocular Depth EstimationJie Long Lee, Gim Hee LeeNeurIPS 2025 · 3 citations
