SemARFlow: Injecting Semantics into Unsupervised Optical Flow Estimation for Autonomous Driving
Shuai Yuan, Shuzhi Yu, Hannah Kim, Carlo Tomasi
Abstract
Unsupervised optical flow estimation is especially hard near occlusions and motion boundaries and in low-texture regions. We show that additional information such as semantics and domain knowledge can help better constrain this problem. We introduce SemARFlow, an unsupervised optical flow network designed for autonomous driving data that takes estimated semantic segmentation masks as additional inputs. This additional information is injected into the encoder and into a learned upsampler that refines the flow output. In addition, a simple yet effective semantic augmentation module provides selfsupervision when learning flow and its boundaries for vehicles, poles, and sky. Together, these injections of semantic information improve the KITTI-2015 optical flow test error rate from 11.80% to 8.38%. We also show visible improvements around object boundaries as well as a greater ability to generalize across datasets. Code is available at https://github.com/duke-vision/ semantic-unsup-flow-release .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext edb14637-bf50-4cf3-865c-5cd6851122b2Cited by top-tier papers5
- UnSAMFlow: Unsupervised Optical Flow Guided by Segment Anything ModelShuai Yuan, Lei Luo, Zhuo Hui, Can Pu et al.CVPR 2024 · 7 citations
- M2Flow: A Motion Information Fusion Framework for Enhanced Unsupervised Optical Flow Estimation in Autonomous DrivingXunpei Sun, Gang Chen, Zuoxun HouAAAI 2025 · 4 citations
- PriOr-Flow: Enhancing Primitive Panoramic Optical Flow with Orthogonal ViewLongliang Liu, Miaojie Feng, Junda Cheng, Jijun Xiang et al.ICCV 2025
- U^2Flow: Uncertainty-Aware Unsupervised Optical Flow EstimationXunpei Sun, Wenwei Lin, Yi Chang, Gang ChenCVPR 2026
- TimeTracker: Event-based Continuous Point Tracking for Video Frame Interpolation with Non-linear MotionHaoyue Liu, Jinghan Xu, Yi Chang, Hanyu Zhou et al.CVPR 2025
Builds on17
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li et al.ICCV 2021 · 402 citations
Related papers
- Imposing Consistency for Optical Flow EstimationJisoo Jeong, Jamie Menjay Lin, Fatih Porikli, Nojun KwakCVPR 2022 · 40 citations
- DistractFlow: Improving Optical Flow Estimation via Realistic Distractions and Pseudo-LabelingJisoo Jeong, Hong Cai, Risheek Garrepalli, Fatih PorikliCVPR 2023
- Learning by Analogy: Reliable Supervision From Transformations for Unsupervised Optical Flow EstimationLiang Liu, Jiangning Zhang, Ruifei He, Yong Liu et al.CVPR 2020
- Flow2Stereo: Effective Self-Supervised Learning of Optical Flow and Stereo MatchingPengpeng Liu, Irwin King, Michael R. Lyu, Jia XuCVPR 2020
- SMURF: Self-Teaching Multi-Frame Unsupervised RAFT With Full-Image WarpingAustin Stone, Daniel Maurer, Alper Ayvaci, Anelia Angelova et al.CVPR 2021
