FlowDiffuser: Advancing Optical Flow Estimation with Diffusion Models
Ao Luo, Xin Li, Fan Yang, Jiangyu Liu, Haoqiang Fan, Shuaicheng Liu
摘要
Optical flow estimation, a process of predicting pixel-wise displacement between consecutive frames, has commonly been approached as a regression task in the age of deep learning. Despite notable advancements, this de facto paradigm unfortunately falls short in generalization performance when trained on synthetic or constrained data. Pioneering a paradigm shift, we reformulate optical flow estimation as a conditional flow generation challenge, unveiling FlowDiffuser — a new family of optical flow models that could have stronger learning and generalization capabilities. FlowDiffuser estimates optical flow through a ‘noise-to-flow’ strategy, progressively eliminating noise from randomly generated flows conditioned on the provided pairs. To optimize accuracy and efficiency, our FlowDiffuser incorporates a novel Conditional Recurrent Denoising Decoder (Conditional-RDD), streamlining the flow estimation process. It incorporates a unique Hidden State Denoising (HSD) paradigm, effectively leveraging the information from previous time steps. Moreover, FlowDiffuser can be easily integrated into existing flow networks, leading to significant improvements in performance metrics compared to conventional implementations. Experiments on challenging benchmarks, including Sintel and KITTI, demonstrate the effectiveness of our FlowDiffuser with superior performance to existing state-of-the-art models. Code is available at https://github.com/LA30/FlowDiffuser.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Learning to See in the Extremely DarkHai Jiang, Binhao Guan, Zhen Liu, Xiaohong Liu 等ICCV 2025 · 被引用 10 次
- ISPDiffuser: Learning RAW-to-sRGB Mappings with Texture-Aware Diffusion Models and Histogram-Guided Color ConsistencyYang Ren, Hai Jiang, Menglong Yang, Wei Li 等AAAI 2025 · 被引用 7 次
- MIORe & VAR-MIORe: Benchmarks to Push the Boundaries of RestorationGeorge Ciubotariu, Zhuyun Zhou, Zongwei Wu, Radu TimofteICCV 2025 · 被引用 2 次
- Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth EstimationXinhao Cai, Gensheng Pei, Zeren Sun, Yazhou Yao 等CVPR 2026 · 被引用 2 次
- DMAligner: Enhancing Image Alignment via Diffusion Model Based View SynthesisXinglong Luo, Ao Luo, Zhengning Wang, Yueqi Yang 等CVPR 2026 · 被引用 1 次
它引用的顶会 Paper25
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- DiffusionDet: Diffusion Model for Object DetectionShoufa Chen, Peize Sun, Yibing Song, Ping LuoICCV 2023 · 被引用 715 次
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li 等ICCV 2021 · 被引用 402 次
相关 Paper
- FlowFM: Advancing Dark Optical Flow Estimation with Flow MatchingFengyuan Zuo, Haiyan Jin, Yuanlin Zhang, Zhaolin Xiao 等CVPR 2026
- The Surprising Effectiveness of Diffusion Models for Optical Flow and Monocular Depth EstimationSaurabh Saxena, Charles Herrmann, Junhwa Hur, Abhishek Kar 等NeurIPS 2023 · 被引用 160 次
- TransFlow: Transformer as Flow LearnerYawen Lu, Qifan Wang, Siqi Ma, Tong Geng 等CVPR 2023
- DistractFlow: Improving Optical Flow Estimation via Realistic Distractions and Pseudo-LabelingJisoo Jeong, Hong Cai, Risheek Garrepalli, Fatih PorikliCVPR 2023
- CRAFT: Cross-Attentional Flow Transformer for Robust Optical FlowXiuchao Sui, Shaohua Li, Xue Geng, Yan Wu 等CVPR 2022 · 被引用 114 次
