RefTeacher: A Strong Baseline for Semi-Supervised Referring Expression Comprehension
Jiamu Sun, Gen Luo, Yiyi Zhou, Xiaoshuai Sun, Guannan Jiang, Zhiyu Wang, Rongrong Ji
摘要
Referring expression comprehension (REC) often requires a large number of instance-level annotations for fully supervised learning, which are laborious and expensive. In this paper, we present the first attempt of semi-supervised learning for REC and propose a strong baseline method called RefTeacher. Inspired by the recent progress in computer vision, RefTeacher adopts a teacher-student learning paradigm, where the teacher REC network predicts pseudolabels for optimizing the student one. This paradigm allows REC models to exploit massive unlabeled data based on a small fraction of labeled. In particular, we also identify two key challenges in semi-supervised REC, namely, sparse supervision signals and worse pseudo-label noise. To address these issues, we equip RefTeacher with two novel designs called Attention-based Imitation Learning (AIL) and Adaptive Pseudo-label Weighting (APW). AIL can help the student network imitate the recognition behaviors of the teacher, thereby obtaining sufficient supervision signals. APW can help the model adaptively adjust the contributions of pseudo-labels with varying qualities, thus avoiding confirmation bias. To validate RefTeacher, we conduct extensive experiments on three REC benchmark datasets. Experimental results show that RefTeacher obtains obvious gains over the fully supervised methods. More importantly, using only 10% labeled data, our approach allows the model to achieve near 100% fully supervised performance, e.g., only -2.78% on RefCOCO. Project: https://refteacher.github.io/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- SAM as the Guide: Mastering Pseudo-Label Refinement in Semi-Supervised Referring Expression SegmentationDanni Yang, Jiayi Ji, Yiwei Ma, Tianyu Guo 等ICML 2024 · 被引用 19 次
- Visual Grounding with Attention-Driven Constraint BalancingWeitai Kang, Luowei Zhou, Junyi Wu, Changchang Sun 等ACM MM 2025 · 被引用 1 次
- WeakMCN: Multi-task Collaborative Network for Weakly Supervised Referring Expression Comprehension and SegmentationSilin Cheng, Yang Liu, Xinwei He, Sébastien Ourselin 等CVPR 2025
- Audio-Visual Segmentation via Unlabeled Frame ExploitationJinxiang Liu, Yikun Liu, Fei Zhang, Chen Ju 等CVPR 2024
- Learning to Label: A Reinforced Self-Evolving Framework for Semi-supervised Referring Expression SegmentationRunlong Cao, Ying Zang, Chuanwei Zhou, Tianrun Chen 等ICML 2026
它引用的顶会 Paper22
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong 等NeurIPS 2020 · 被引用 2,774 次
- FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo LabelingBowen Zhang, Yidong Wang, Wenxin Hou, Hao Wu 等NeurIPS 2021 · 被引用 1,389 次
- MDETR - Modulated Detection for End-to-End Multi-Modal UnderstandingAishwarya Kamath, Mannat Singh, Yann LeCun, Gabriel Synnaeve 等ICCV 2021 · 被引用 1,114 次
- Unbiased Teacher for Semi-Supervised Object DetectionYen-Cheng Liu, Chih-Yao Ma, Zijian He, Chia-Wen Kuo 等ICLR 2021 · 被引用 603 次
相关 Paper
- RefCLIP: A Universal Teacher for Weakly Supervised Referring Expression ComprehensionLei Jin, Gen Luo, Yiyi Zhou, Xiaoshuai Sun 等CVPR 2023
- Learning to Segment Every Referring Object Point by PointMengxue Qu, Yu Wu, Yunchao Wei, Wu Liu 等CVPR 2023
- De-biased Teacher: Rethinking IoU Matching for Semi-supervised Object DetectionKuo Wang, Jingyu Zhuang, Guanbin Li, Chaowei Fang 等AAAI 2023 · 被引用 16 次
- Whether you can locate or not? Interactive Referring Expression GenerationFulong Ye, Yuxing Long, Fangxiang Feng, Xiaojie WangACM MM 2023 · 被引用 6 次
- Label Matching Semi-Supervised Object DetectionBinbin Chen, Weijie Chen, Shicai Yang, Yunyi Xuan 等CVPR 2022 · 被引用 87 次
