Vision Transformers are Good Mask Auto-Labelers
Shiyi Lan, Xitong Yang, Zhiding Yu, Zuxuan Wu, José M. Álvarez, Anima Anandkumar
2023Year
9Top-tier citations
Abstract
https://github.com/NVlabs/mask-auto-labeler Figure 1. Examples of mask pseudo-labels generated by Mask Auto-Labeler on COCO. Only human-annotated bounding boxes are used as supervision during training to obtain these results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2657f271-7a05-42a4-a21e-79fc92e5255dCited by top-tier papers9
- A Simple Framework for Open-Vocabulary Segmentation and DetectionHao Zhang, Feng Li, Xueyan Zou, Shilong Liu et al.ICCV 2023 · 241 citations
- Label-efficient Segmentation via Affinity PropagationWentong Li, Yuqian Yuan, Song Wang, Wenyu Liu et al.NeurIPS 2023 · 11 citations
- GaPro: Box-Supervised 3D Point Cloud Instance Segmentation Using Gaussian Processes as Pseudo LabelersTuan Duc Ngo, Binh-Son Hua, Khoi NguyenICCV 2023 · 9 citations
- MonoBox: Tightness-Free Box-Supervised Polyp Segmentation Using Monotonicity ConstraintQiang Hu, Zhenyu Yi, Ying Zhou, Fan Huang et al.AAAI 2025 · 8 citations
- Advancing Visual Large Language Model for Multi-Granular Versatile PerceptionWentao Xiang, Haoxian Tan, Yujie Zhong, Cong Wei et al.ICCV 2025 · 1 citation
Builds on23
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
Related papers
- Open-Vocabulary Instance Segmentation via Robust Cross-Modal Pseudo-LabelingDat Huynh, Jason Kuen, Zhe Lin, Jiuxiang Gu et al.CVPR 2022 · 78 citations
- Pseudo-label Alignment for Semi-supervised Instance SegmentationJie Hu, Chen Chen, Liujuan Cao, Shengchuan Zhang et al.ICCV 2023 · 31 citations
- BBAM: Bounding Box Attribution Map for Weakly Supervised Semantic and Instance SegmentationJungbeom Lee, Jihun Yi, Chaehun Shin, Sungroh YoonCVPR 2021
- SparseDet: Improving Sparsely Annotated Object Detection with Pseudo-positive MiningSaksham Suri, Sai Saketh Rambhatla, Rama Chellappa, Abhinav ShrivastavaICCV 2023 · 19 citations
- Unsupervised Universal Image SegmentationDantong Niu, Xudong Wang, Xinyang Han, Long Lian et al.CVPR 2024 · 29 citations
