FlowSeek: Optical Flow Made Easier with Depth Foundation Models and Motion Bases
Matteo Poggi, Fabio Tosi
Abstract
We present FlowSeek, a novel framework for optical flow requiring minimal hardware resources for training. FlowSeek marries the latest advances on the design space of optical flow networks with cutting-edge single-image depth foundation models and classical low-dimensional motion parametrization, implementing a compact, yet accurate architecture. FlowSeek is trained on a single consumer-grade GPU, a hardware budget about lower compared to most recent methods, and still achieves superior cross-dataset generalization on Sintel Final and KITTI, with a relative improvement of 10 and 15% over the previous state-of-the-art SEA-RAFT, as well as on Spring and LayeredFlow datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 895714ce-3f1b-4bce-9f1e-b3feae6808a6Cited by top-tier papers4
- CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image RegistrationXuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li et al.CVPR 2026 · 4 citations
- Optical Flow Matching: Reframing Optical Flow as Continuous Transport DynamicsAo Luo, Xin Li, Fan Yang, Yuezun Li et al.CVPR 2026
- ARFlow: Auto-regressive Optical Flow Estimation for Arbitrary-Length Videos via Progressive Next-Frame ForecastingJiuming Liu, Mengmeng Liu, Siting Zhu, Yunpeng Zhang et al.ICLR 2026
- High Resolution Neural Video Coding with Bi-directional Confidence-Guided Reference Information ModelingFeng Ye, Kai Zhang, Li Zhang, Chuanmin JiaCVPR 2026
Builds on43
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao et al.NeurIPS 2024 · 2,305 citations
Related papers
- MEMFOF: High-Resolution Training for Memory-Efficient Multi-Frame Optical Flow EstimationVladislav Bargatin, Egor Chistov, Alexander Yakovenko, Dmitriy S. VatolinICCV 2025 · 11 citations
- WAFT: Warping-Alone Field Transforms for Optical FlowYihan Wang, Jia DengICLR 2026 · 36 citations
- ScopeFlow: Dynamic Scene Scoping for Optical FlowAviram Bar-Haim, Lior WolfCVPR 2020
- Learning Optical Flow From Still ImagesFilippo Aleotti, Matteo Poggi, Stefano MattocciaCVPR 2021
- SMURF: Self-Teaching Multi-Frame Unsupervised RAFT With Full-Image WarpingAustin Stone, Daniel Maurer, Alper Ayvaci, Anelia Angelova et al.CVPR 2021
