FLOAT: Factorized Learning of Object Attributes for Improved Multi-object Multi-part Scene Parsing
Rishubh Singh, Pranav Gupta, Pradeep Shenoy, Ravikiran Sarvadevabhatla
摘要
Multi-object multi-part scene parsing is a challenging task which requires detecting multiple object classes in a scene and segmenting the semantic parts within each object. In this paper, we propose FLOAT, a factorized label space framework for scalable multi-object multi-part parsing. Our framework involves independent dense prediction of object category and part attributes which increases scalability and reduces task complexity compared to the monolithic label space counterpart. In addition, we propose an inference-time ‘zoom’ refinement technique which significantly improves segmentation quality, especially for smaller objects/parts. Compared to state of the art, FLOAT obtains an absolute improvement of 2.0% for mean IOU (mIOU) and 4.8% for segmentation quality IOU (sqIOU) on the Pascal-Part-58 dataset. For the larger Pascal-Part-108 dataset, the improvements are 2.1% for mIOU and 3.9% for sqIOU. We incorporate previously excluded part attributes and other minor parts of the Pascal-Part dataset to create the most comprehensive and challenging version which we dub Pascal-Part-201. FLOAT obtains improvements of 8.6% for mIOU and 7.5% for sqIOU on the new dataset, demonstrating its parsing effectiveness across a challenging diversity of objects and parts. The code and datasets are available at floatseg.github.io.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- LogicSeg: Parsing Visual Semantics with Neural Logic Learning and ReasoningLiulei Li, Wenguan Wang, Yang YiICCV 2023 · 被引用 52 次
- MaskPrompt: Open-Vocabulary Affordance Segmentation with Object Shape Mask PromptsDongpan Chen, Dehui Kong, Jinghua Li, Baocai YinAAAI 2025 · 被引用 5 次
- Towards Open-World Segmentation of PartsTai-Yu Pan, Qing Liu, Wei-Lun Chao, Brian L. PriceCVPR 2023
- CALICO: Part-Focused Semantic Co-Segmentation with Large Vision-Language ModelsKiet A. Nguyen, Adheesh Sunil Juvekar, Tianjiao Yu, Muntasir Wahed 等CVPR 2025
- Part-Aware Panoptic SegmentationDaan de Geus, Panagiotis Meletis, Chenyang Lu, Xiaoxiao Wen 等CVPR 2021
它引用的顶会 Paper8
- Shapeglot: Learning Language for Shape DifferentiationPanos Achlioptas, Leonidas J. Guibas, Noah D. Goodman, Judy Fan 等ICCV 2019 · 被引用 86 次
- Composite Shape Modeling via Latent Space FactorizationAnastasia Dubrovina, Fei Xia, Panos Achlioptas, Mira Shalah 等ICCV 2019 · 被引用 66 次
- Multi-Class Part Parsing With Joint Boundary-Semantic AwarenessYifan Zhao, Jia Li, Yu Zhang, Yonghong TianICCV 2019 · 被引用 63 次
- PTR: A Benchmark for Part-based Conceptual, Relational, and Physical ReasoningYining Hong, Li Yi, Josh Tenenbaum, Antonio Torralba 等NeurIPS 2021 · 被引用 46 次
- Hybrid Resolution Network Using Edge Guided Region Mutual Information Loss for Human ParsingYunan Liu, Liang Zhao, Shanshan Zhang, Jian YangACM MM 2020 · 被引用 21 次
相关 Paper
- Going Denser with Open-Vocabulary Part SegmentationPeize Sun, Shoufa Chen, Chenchen Zhu, Fanyi Xiao 等ICCV 2023 · 被引用 83 次
- Compositor: Bottom-Up Clustering and Compositing for Robust Part and Object SegmentationJu He, Jieneng Chen, Ming-Xian Lin, Qihang Yu 等CVPR 2023
- Understanding Multi-Granularity for Open-Vocabulary Part SegmentationJiho Choi, Seonho Lee, Seungho Lee, Minhyun Lee 等NeurIPS 2024 · 被引用 7 次
- Dynamic Multi-Scale Filters for Semantic SegmentationJunjun He, Zhongying Deng, Yu QiaoICCV 2019 · 被引用 287 次
- PACO: Parts and Attributes of Common ObjectsVignesh Ramanathan, Anmol Kalia, Vladan Petrovic, Yi Wen 等CVPR 2023
