Boosting Robustness of Image Matting with Context Assembling and Strong Data Augmentation
Yutong Dai, Brian L. Price, He Zhang, Chunhua Shen
摘要
Deep image matting methods have achieved increasingly better results on benchmarks (e.g., Composition-1k/alphamatting.com). However, the robustness, including robustness to trimaps and generalization to images from different domains, is still underexplored. Although some works propose to either refine the trimaps or adapt the algorithms to real-world images via extra data augmentation, none of them has taken both into consideration, not to mention the significant performance deterioration on benchmarks while using those data augmentation. To fill this gap, we propose an image matting method which achieves higher robustness (RMat) via multilevel context assembling and strong data augmentation targeting matting. Specifically, we first build a strong matting framework by modeling ample global information with transformer blocks in the encoder, and focusing on details in combination with convolution layers as well as a low-level feature assembling attention block in the decoder. Then, based on this strong baseline, we analyze current data augmentation and explore simple but effective strong data augmentation to boost the baseline model and contribute a more generalizable matting method. Compared with previous methods, the proposed method not only achieves state-of-the-art results on the Composition-1k benchmark (11 % improvement on SAD and 27% improvement on Grad) with smaller model size, but also shows more robust generalization results on other benchmarks, on real-world images, and also on varying coarse-to-fine trimaps with our extensive experiments. <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup> <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup> This work was in part done when YD was an intern at Adobe and CS was with The University of Adelaide. CS is the corresponding author. Project page: https://dongdong93.github.io/RMat/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Revisiting Context Aggregation for Image MattingQinglin Liu, Xiaoqian Lv, Quanling Meng, Zonglin Li 等ICML 2024 · 被引用 6 次
- Infusing Definiteness into Randomness: Rethinking Composition Styles for Deep Image MattingZixuan Ye, Yutong Dai, Chaoyi Hong, Zhiguo Cao 等AAAI 2023 · 被引用 2 次
- Memory Efficient Matting with Adaptive Token RoutingYiheng Lin, Yihan Hu, Chenyi Zhang, Ting Liu 等AAAI 2025 · 被引用 1 次
- Polar Matte: Fully Computational Ground-Truth-Quality Alpha Matte Extraction for Images and Video using Polarized Screen MattingKenji Enomoto, T. J. Rhodes, Brian L. Price, Gavin MillerCVPR 2024
- Segment and Matte Anything in a Unified ModelZezhong Fan, Xiaohan Li, Topojoy Biswas, Kaushiki Nag 等AAAI 2026
它引用的顶会 Paper17
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan 等ICCV 2021 · 被引用 4,909 次
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li 等AAAI 2020 · 被引用 4,134 次
- ConViT: Improving Vision Transformers with Soft Convolutional Inductive BiasesStéphane d'Ascoli, Hugo Touvron, Matthew L. Leavitt, Ari S. Morcos 等ICML 2021 · 被引用 1,021 次
相关 Paper
- Disentangled Image MattingShaofan Cai, Xiaoshuai Zhang, Haoqiang Fan, Haibin Huang 等ICCV 2019 · 被引用 127 次
- High-Resolution Deep Image MattingHaichao Yu, Ning Xu, Zilong Huang, Yuqian Zhou 等AAAI 2021 · 被引用 61 次
- Mask-Guided Matting in the WildKwanyong Park, Sanghyun Woo, Seoung Wug Oh, In So Kweon 等CVPR 2023
- Tripartite Information Mining and Integration for Image MattingYuhao Liu, Jiake Xie, Xiao Shi, Yu Qiao 等ICCV 2021 · 被引用 66 次
- Natural Image Matting via Guided Contextual AttentionYaoyi Li, Hongtao LuAAAI 2020 · 被引用 189 次
