Injecting Frame-Event Complementary Fusion into Diffusion for Optical Flow in Challenging Scenes
Haonan Wang, Hanyu Zhou, Haoyue Liu, Luxin Yan
摘要
Optical flow estimation has achieved promising results in conventional scenes but faces challenges in high-speed and low-light scenes, which suffer from motion blur and insufficient illumination. These conditions lead to weakened texture and amplified noise and deteriorate the appearance saturation and boundary completeness of frame cameras, which are necessary for motion feature matching. In degraded scenes, the frame camera provides dense appearance saturation but sparse boundary completeness due to its long imaging time and low dynamic range. In contrast, the event camera offers sparse appearance saturation, while its short imaging time and high dynamic range gives rise to dense boundary completeness. Traditionally, existing methods utilize feature fusion or domain adaptation to introduce event to improve boundary completeness. However, the appearance features are still deteriorated, which severely affects the mostly adopted discriminative models that learn the mapping from visual features to motion fields and generative models that generate motion fields based on given visual features. So we introduce diffusion models that learn the mapping from noising flow to clear flow, which is not affected by the deteriorated visual features. Therefore, we propose a novel optical flow estimation framework Diff-ABFlow based on diffusion models with frame-event appearance-boundary fusion. Inspired by the appearanceboundary complementarity of frame and event, we propose an Attention-Guided Appearance-Boundary Fusion module to fuse frame and event. Based on diffusion models, we propose a Multi-Condition Iterative Denoising Decoder. Our proposed method can effectively utilize the respective advantages of frame and event, and shows great robustness to degraded input. In addition, we propose a dual-modal optical flow dataset for generalization experiments. Extensive experiments have verified the superiority of our proposed method. The code is released at https://github.com/Haonan-Wang-aurora/Diff-ABFlow.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper19
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 被引用 11,724 次
- Label-Efficient Semantic Segmentation with Diffusion ModelsDmitry Baranchuk, Andrey Voynov, Ivan Rubachev, Valentin Khrulkov 等ICLR 2022 · 被引用 700 次
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li 等ICCV 2021 · 被引用 402 次
- GMFlow: Learning Optical Flow via Global MatchingHaofei Xu, Jing Zhang, Jianfei Cai, Hamid Rezatofighi 等CVPR 2022 · 被引用 353 次
- The Surprising Effectiveness of Diffusion Models for Optical Flow and Monocular Depth EstimationSaurabh Saxena, Charles Herrmann, Junhwa Hur, Abhishek Kar 等NeurIPS 2023 · 被引用 160 次
相关 Paper
- Bridge Frame and Event: Common Spatiotemporal Fusion for High-Dynamic Scene Optical FlowHanyu Zhou, Haonan Wang, Haoyue Liu, Yuxing Duan 等CVPR 2025
- NEC-Diff: Noise-Robust Event-RAW Complementary Diffusion for Seeing Motion in Extreme DarknessHaoyue Liu, Jinghan Xu, Luxin Feng, Hanyu Zhou 等CVPR 2026
- ESEG: Event-Based Segmentation Boosted by Explicit Edge-Semantic GuidanceYucheng Zhao, Gengyu Lyu, Ke Li, Zihao Wang 等AAAI 2025 · 被引用 8 次
- Event-based Motion Deblurring with Modality-Aware Decomposition and RecompositionWen Yang, Jinjian Wu, Leida Li, Weisheng Dong 等ACM MM 2023 · 被引用 14 次
- Learning for Motion Deblurring with Hybrid Frames and EventsWen Yang, Jinjian Wu, Jupo Ma, Leida Li 等ACM MM 2022 · 被引用 17 次
