VSCD: Video-based Scene Change Detection in Unaligned Scenes
Jiae Yoon, Ue-Hwan Kim
摘要
Detecting what has changed in an environment is essential for long-term autonomy, yet most change detection settings assume fixed viewpoints, mild misalignment, or only a few changed objects. We introduce Video-based Scene Change Detection (VSCD), which predicts a pixel-wise change mask for each query frame, given a reference and a query RGB video of the same indoor space recorded at different times under unconstrained camera motion. The two videos are not temporally synchronized, and many object instances may appear or disappear. To study this setting, we build a large-scale benchmark with over 1.1 million frames annotated with pixel-accurate change masks, together with a real-world test set for evaluating transfer beyond simulation. We propose a query-centric multi-reference model that learns temporal matching implicitly from change-mask supervision, aligns candidate reference features to the query via local patch correspondence, and fuses per-candidate change features using frame-level and patch-level confidence before decoding a high-resolution mask once per frame. Our approach achieves state-of-the-art performance against strong image- and video-based baselines, and we validate its real-world impact by deploying it on a mobile robot for two downstream applications—visual surveillance and object incremental learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch 等ICML 2023 · 被引用 2,601 次
- Video Similarity and Alignment Learning on Partial Video Copy DetectionZhen Han, Xiangteng He, Mingqian Tang, Yiliang LvACM MM 2021 · 被引用 34 次
- TransVCL: Attention-Enhanced Video Copy Localization Network with Flexible SupervisionSifeng He, Yue He, Minlong Lu, Chen Jiang 等AAAI 2023 · 被引用 26 次
- Zero-Shot Scene Change DetectionKyusik Cho, Dong Yeop Kim, Euntai KimAAAI 2025 · 被引用 10 次
相关 Paper
- Dual Task Learning by Leveraging Both Dense Correspondence and Mis-Correspondence for Robust Change Detection With Imperfect MatchesJin-Man Park, Ue-Hwan Kim, Seon-Hoon Lee, Jong-Hwan KimCVPR 2022 · 被引用 13 次
- Towards Generalizable Scene Change DetectionJae-Woo Kim, Ue-Hwan KimCVPR 2025
- Changes in Real Time: Online Scene Change Detection with Multi-View FusionChamuditha Jayanga Galappaththige, Jason Lai, Lloyd Windrim, Donald G. Dansereau 等CVPR 2026 · 被引用 4 次
- 3D Video Object Detection with Learnable Object-Centric Global OptimizationJiawei He, Yuntao Chen, Naiyan Wang, Zhaoxiang ZhangCVPR 2023
- Multi-View Pose-Agnostic Change Localization with Zero LabelsChamuditha Jayanga Galappaththige, Jason Lai, Lloyd Windrim, Donald G. Dansereau 等CVPR 2025
