MUVA: A New Large-Scale Benchmark for Multi-view Amodal Instance Segmentation in the Shopping Scenario
Zhixuan Li, Weining Ye, Juan R. Terven, Zachary Bennett, Ying Zheng, Tingting Jiang, Tiejun Huang
摘要
Amodal Instance Segmentation (AIS) endeavors to accurately deduce complete object shapes that are partially or fully occluded. However, the inherent ill-posed nature of single-view datasets poses challenges in determining occluded shapes. A multi-view framework may help alleviate this problem, as humans often adjust their perspective when encountering occluded objects. At present, this approach has not yet been explored by existing methods and datasets. To bridge this gap, we propose a new task called Multi-view Amodal Instance Segmentation (MAIS) and introduce the MUVA dataset, the first MUlti-View AIS dataset that takes the shopping scenario as instantiation. MUVA provides comprehensive annotations, including multi-view amodal/visible segmentation masks, 3D models, and depth maps, making it the largest image-level AIS dataset in terms of both the number of images and instances. Additionally, we propose a new method for aggregating representative features across different instances and views, which demonstrates promising results in accurately predicting occluded objects from one viewpoint by leveraging information from other viewpoints. Besides, we also demonstrate that MUVA can benefit the AIS task in real-world scenarios. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- MonoMAE: Enhancing Monocular 3D Detection through Depth-Aware Masked AutoencodersXueying Jiang, Sheng Jin, Xiaoqin Zhang, Ling Shao 等NeurIPS 2024 · 被引用 31 次
- Amodal Ground Truth and Completion in the WildGuanqi Zhan, Chuanxia Zheng, Weidi Xie, Andrew ZissermanCVPR 2024 · 被引用 23 次
- Unlocking Constraints: Source-Free Occlusion-Aware Seamless SegmentationYihong Cao, Jiaming Zhang, Xu Zheng, Hao Shi 等ICCV 2025 · 被引用 4 次
- SynergyAmodal: Deocclude Anything with Text ControlXinyang Li, Chengjie Yi, Jiawei Lai, Mingbao Lin 等ACM MM 2025 · 被引用 3 次
- Stable Diffusion-Based Approach for Human De-OcclusionSeung Young Noh, Ju Yong ChangACM MM 2025
它引用的顶会 Paper12
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi 等ICCV 2019 · 被引用 3,348 次
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 被引用 1,416 次
- Domain Generalization with MixStyleKaiyang Zhou, Yongxin Yang, Yu Qiao, Tao XiangICLR 2021 · 被引用 986 次
- E2EC: An End-to-End Contour-based Method for High-Quality High-Speed Instance SegmentationTao Zhang, Shiqing Wei, Shunping JiCVPR 2022 · 被引用 110 次
- Amodal Segmentation Based on Visible Region Segmentation and Shape PriorYuting Xiao, Yanyu Xu, Ziming Zhong, Weixin Luo 等AAAI 2021 · 被引用 76 次
相关 Paper
- Amodal Instance Segmentation with IRAIS Dataset for Sim-to-Real TransferBidong Chen, Lingui LiICML 2026
- Unveiling the Invisible: Reasoning Complex Occlusions Amodally with AURAZhixuan Li, Hyunse Yoon, Sanghoon Lee, Weisi LinICCV 2025
- Segment Anything, Even OccludedWei-En Tai, Yu-Lin Shih, Cheng Sun, Yu-Chiang Frank Wang 等CVPR 2025
- Amodal Scene Analysis via Holistic Occlusion Relation Inference and Generative Mask CompletionBowen Zhang, Qing Liu, Jianming Zhang, Yilin Wang 等AAAI 2024 · 被引用 4 次
- Variational Amodal Object CompletionHuan Ling, David Acuna, Karsten Kreis, Seung Wook Kim 等NeurIPS 2020 · 被引用 56 次
