ZOOM: Learning Video Mirror Detection with Extremely-Weak Supervision
Ke Xu, Tsun Wai Siu, Rynson W. H. Lau
Abstract
Mirror detection is an active research topic in computer vision. However, all existing mirror detectors learn mirror representations from large-scale pixel-wise datasets, which are tedious and expensive to obtain. Although weakly-supervised learning has been widely explored in related topics, we note that popular weak supervision signals (e.g., bounding boxes, scribbles, points) still require some efforts from the user to locate the target objects, with a strong assumption that the images to annotate always contain the target objects. Such an assumption may result in the over-segmentation of mirrors. Our key idea of this work is that the existence of mirrors over a time period may serve as a weak supervision to train a mirror detector, for two reasons. First, if a network can predict the existence of mirrors, it can essentially locate the mirrors. Second, we observe that the reflected contents of a mirror tend to be similar to those in adjacent frames, but exhibit considerable contrast to regions in far-away frames (e.g., non-mirror frames). To this end, in this paper, we propose ZOOM, the first method to learn robust mirror representations from extremely-weak annotations of per-frame ZerO-One Mirror indicators in videos. The key insight of ZOOM is to model the similarity and contrast (between mirror and non-mirror regions) in temporal variations to locate and segment the mirrors. To this end, we propose a novel fusion strategy to leverage temporal consistency information for mirror localization, and a novel temporal similarity-contrast modeling module for mirror segmentation. We construct a new video mirror dataset for training and evaluation. Experimental results under new and standard metrics show that ZOOM performs favorably against existing fully-supervised mirror detection methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext aec49a78-ae61-4a6a-ad36-7f9382d31a8dCited by top-tier papers2
- Beyond Single Images: Retrieval Self-Augmented Unsupervised Camouflaged Object DetectionJi Du, Xin Wang, Fangwei Hao, Mingyang Yu et al.ICCV 2025 · 2 citations
- Seeing Beyond Illusion: Generalized and Efficient Mirror DetectionMingfeng Zha, Guoqing Wang, Tianyu Li, Wei Dong et al.AAAI 2026
Builds on26
- Rethinking Efficient Lane Detection via Curve ModelingZhengyang Feng, Shaohua Guo, Xin Tan, Ke Xu et al.CVPR 2022 · 204 citations
- Motion Guided Attention for Video Salient Object DetectionHaofeng Li, Guanqi Chen, Guanbin Li, Yizhou YuICCV 2019 · 200 citations
- Pyramid Constrained Self-Attention Network for Fast Video Salient Object DetectionYuchao Gu, Lijuan Wang, Ziqin Wang, Yun Liu et al.AAAI 2020 · 184 citations
- Structure-Consistent Weakly Supervised Salient Object Detection with Local Saliency CoherenceSiyue Yu, Bingfeng Zhang, Jimin Xiao, Eng Gee LimAAAI 2021 · 162 citations
- Unlocking the Potential of Ordinary Classifier: Class-specific Adversarial Erasing Framework for Weakly Supervised Semantic SegmentationHyeokjun Kweon, Sung-Hoon Yoon, Hyeonseong Kim, Daehee Park et al.ICCV 2021 · 151 citations
Related papers
- Weakly-Supervised Mirror Detection via Scribble AnnotationsMingfeng Zha, Yunqiang Pei, Guoqing Wang, Tianyu Li et al.AAAI 2024 · 18 citations
- Where Is My Mirror?Xin Yang, Haiyang Mei, Ke Xu, Xiaopeng Wei et al.ICCV 2019 · 6 citations
- Progressive Mirror DetectionJiaying Lin, Guodong Wang, Rynson W. H. LauCVPR 2020
- Weakly Supervised Video Moment Localization with Contrastive Negative Sample MiningMinghang Zheng, Yanjie Huang, Qingchao Chen, Yang LiuAAAI 2022 · 109 citations
- Learning to Detect Mirrors from Videos via Dual CorrespondencesJiaying Lin, Xin Tan, Rynson W. H. LauCVPR 2023
