Event Camera Data Pre-training
Yan Yang, Liyuan Pan, Liu Liu
摘要
This paper proposes a pre-trained neural network for handling event camera data. Our model is a self-supervised learning framework, and uses paired event camera data and natural RGB images for training. Our method contains three modules connected in a sequence: i) a family of event data augmentations, generating meaningful event images for self-supervised training; ii) a conditional masking strategy to sample informative event patches from event images, encouraging our model to capture the spatial layout of a scene and accelerating training; iii) a contrastive learning approach, enforcing the similarity of embeddings between matching event images, and between paired event and RGB images. An embedding projection loss is proposed to avoid the model collapse when enforcing the event image embedding similarities. A probability distribution alignment loss is proposed to encourage the event image to be consistent with its paired RGB image in the feature space. Transfer learning performance on downstream tasks shows the superiority of our method over state-of-the-art methods. For example, we achieve top-1 accuracy at 64.83% on the N-ImageNet dataset. Our code is available at https://github.com/Yan98/Event-Camera-Data-Pre-training.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Language-driven All-in-one Adverse Weather RemovalHao Yang, Liyuan Pan, Yan Yang, Wei LiangCVPR 2024 · 被引用 28 次
- FlexEvent: Towards Flexible Event-Frame Object Detection at Varying Operational FrequenciesDongyue Lu, Lingdong Kong, Gim Hee Lee, Camille Simon Chane 等NeurIPS 2025 · 被引用 13 次
- Efficient Meshflow and Optical Flow Estimation from Event CamerasXinglong Luo, Ao Luo, Zhengning Wang, Chunyu Lin 等CVPR 2024 · 被引用 10 次
- LEOD: Label-Efficient Object Detection for Event CamerasZiyi Wu, Mathias Gehrig, Qing Lyu, Xudong Liu 等CVPR 2024 · 被引用 10 次
- LDP: Language-driven Dual-Pixel Image Defocus Deblurring NetworkHao Yang, Liyuan Pan, Yan Yang, Richard I. Hartley 等CVPR 2024 · 被引用 9 次
它引用的顶会 Paper16
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
相关 Paper
- Revealing Latent Information: A Physics-inspired Self-supervised Pre-training Framework for Noisy and Sparse EventsLin Zhu, Ruonan Liu, Xiao Wang, Lizhi Wang 等ACM MM 2025 · 被引用 1 次
- CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training FrameworkWentao Wu, Xiao Wang, Chenglong Li, Bo Jiang 等ACM MM 2025 · 被引用 2 次
- Unsupervised Domain Adaptation for Training Event-Based Networks Using Contrastive Learning and Uncorrelated ConditioningDayuan Jian, Mohammad RostamiICCV 2023 · 被引用 22 次
- EZSR: Event-based Zero-Shot RecognitionYan Yang, Liyuan Pan, Dongxu Li, Liu LiuCVPR 2025
- N-ImageNet: Towards Robust, Fine-Grained Object Recognition with Event CamerasJunho Kim, Jaehyeok Bae, Gangin Park, Dongsu Zhang 等ICCV 2021 · 被引用 127 次
