Robust Frame-to-Frame Camera Rotation Estimation in Crowded Scenes
Fabien Delattre, David Dirnfeld, Phat Nguyen, Stephen Scarano, Michael J. Jones, Pedro Miraldo, Erik G. Learned-Miller
Abstract
We present an approach to estimating camera rotation in crowded, real-world scenes from handheld monocular video. While camera rotation estimation is a well-studied problem, no previous methods exhibit both high accuracy and acceptable speed in this setting. Because the setting is not addressed well by other datasets, we provide a new dataset and benchmark, with high-accuracy, rigorously verified ground truth, on 17 video sequences. Methods developed for wide baseline stereo (e.g., 5-point methods) perform poorly on monocular video. On the other hand, methods used in autonomous driving (e.g., SLAM) leverage specific sensor setups, specific motion models, or local optimization strategies (lagging batch processing) and do not generalize well to handheld video. Finally, for dynamic scenes, commonly used robustification techniques like RANSAC require large numbers of iterations, and become prohibitively slow. We introduce a novel generalization of the Hough transform on SO(3) to efficiently and robustly find the camera rotation most compatible with optical flow. Among comparably fast methods, ours reduces error by almost 50% over the next best, and is more accurate than any method, irrespective of speed. This represents a strong new performance point for crowded scenes, an important setting for computer vision. The code and the dataset are available at https://fabiendelattre.com/robustrotation-estimation .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ddceb924-a858-43b5-ad8b-e01394771aadBuilds on3
- DiffPoseNet: Direct Differentiable Camera Pose EstimationChethan M. Parameshwara, Gokul Hari, Cornelia Fermüller, Nitin J. Sanket et al.CVPR 2022 · 35 citations
- Relative Pose from a Calibrated and an Uncalibrated Smartphone ImageYaqing Ding, Daniel Barath, Jian Yang, Zuzana KukelovaCVPR 2022 · 5 citations
- Globally Optimal Relative Pose Estimation With Gravity PriorYaqing Ding, Daniel Barath, Jian Yang, Hui Kong et al.CVPR 2021
Related papers
- Estimating 2D Camera Motion with Hybrid Motion BasisHaipeng Li, Tianhao Zhou, Zhanglei Yang, Yi Wu et al.ICCV 2025
- Robust Consistent Video Depth EstimationJohannes Kopf, Xuejian Rong, Jia-Bin HuangCVPR 2021
- MegaSaM: Accurate, Fast and Robust Structure and Motion from Casual Dynamic VideosZhengqi Li, Richard Tucker, Forrester Cole, Qianqian Wang et al.CVPR 2025
- Princeton365: A Diverse Dataset with Accurate Camera PoseKarhan Kayan, Stamatis Alexandropoulos, Rishabh Jain, Yiming Zuo et al.ICCV 2025
- Real-time Vanishing Point Detector Integrating Under-parameterized RANSAC and Hough TransformJianping Wu, Liang Zhang, Ye Liu, Ke ChenICCV 2021 · 17 citations
