MAOAM: Unified Object and Material Selection with Vision-Language Models
Jaden Park, Valentin Deschaintre, Jason Kuen, Kangning Liu, Iliyan Georgiev, Krishna Kumar Singh, Yong Jae Lee, Michael Fischer
摘要
Selection is a core operation in interactive image editing, enabling tasks such as composition or manipulation. To be practically useful, a user should be able to specify and disambiguate the desired selection region through either text- or click-based interactions, and the system should support selecting not only objects but also other criteria, such as materials. Material-based selection can be particularly valuable for tasks like re-texturing surfaces or consistently editing all instances of a specific material in a scene. However, existing vision–language-model (VLM) based selection methods are largely object-centric and typically support only a single interaction modality, limiting their applicability in real editing workflows.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Materialistic: Selecting Similar Materials in ImagesPrafull Sharma, Julien Philip, Michaël Gharbi, Bill Freeman 等SIGGRAPH 2023 · 被引用 23 次
- PBR3DGen: A VLM-Guided Mesh Generation with High-Quality PBR TextureXiaokang Wei, Bowen Zhang, Xianghui Yang, Yuxuan Wang 等AAAI 2026 · 被引用 1 次
- Alterbute: Editing Intrinsic Attributes of Objects in ImagesTal Reiss, Daniel Winter, Matan Cohen, Alex Rav-Acha 等ICML 2026
- Unifying Automatic and Interactive Matting with Pretrained ViTsZixuan Ye, Wenze Liu, He Guo, Yujia Liang 等CVPR 2024 · 被引用 7 次
- SINE: Semantic-driven Image-based NeRF Editing with Prior-guided Editing FieldChong Bao, Yinda Zhang, Bangbang Yang, Tianxing Fan 等CVPR 2023
