FLOAT: Factorized Learning of Object Attributes for Improved Multi-object Multi-part Scene Parsing
Rishubh Singh, Pranav Gupta, Pradeep Shenoy, Ravikiran Sarvadevabhatla
Abstract
Multi-object multi-part scene parsing is a challenging task which requires detecting multiple object classes in a scene and segmenting the semantic parts within each object. In this paper, we propose FLOAT, a factorized label space framework for scalable multi-object multi-part parsing. Our framework involves independent dense prediction of object category and part attributes which increases scalability and reduces task complexity compared to the monolithic label space counterpart. In addition, we propose an inference-time ‘zoom’ refinement technique which significantly improves segmentation quality, especially for smaller objects/parts. Compared to state of the art, FLOAT obtains an absolute improvement of 2.0% for mean IOU (mIOU) and 4.8% for segmentation quality IOU (sqIOU) on the Pascal-Part-58 dataset. For the larger Pascal-Part-108 dataset, the improvements are 2.1% for mIOU and 3.9% for sqIOU. We incorporate previously excluded part attributes and other minor parts of the Pascal-Part dataset to create the most comprehensive and challenging version which we dub Pascal-Part-201. FLOAT obtains improvements of 8.6% for mIOU and 7.5% for sqIOU on the new dataset, demonstrating its parsing effectiveness across a challenging diversity of objects and parts. The code and datasets are available at floatseg.github.io.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a554b29f-6e89-4143-becf-186b35d20070Cited by top-tier papers8
- LogicSeg: Parsing Visual Semantics with Neural Logic Learning and ReasoningLiulei Li, Wenguan Wang, Yang YiICCV 2023 · 52 citations
- MaskPrompt: Open-Vocabulary Affordance Segmentation with Object Shape Mask PromptsDongpan Chen, Dehui Kong, Jinghua Li, Baocai YinAAAI 2025 · 5 citations
- Towards Open-World Segmentation of PartsTai-Yu Pan, Qing Liu, Wei-Lun Chao, Brian L. PriceCVPR 2023
- CALICO: Part-Focused Semantic Co-Segmentation with Large Vision-Language ModelsKiet A. Nguyen, Adheesh Sunil Juvekar, Tianjiao Yu, Muntasir Wahed et al.CVPR 2025
- Part-Aware Panoptic SegmentationDaan de Geus, Panagiotis Meletis, Chenyang Lu, Xiaoxiao Wen et al.CVPR 2021
Builds on8
- Shapeglot: Learning Language for Shape DifferentiationPanos Achlioptas, Leonidas J. Guibas, Noah D. Goodman, Judy Fan et al.ICCV 2019 · 86 citations
- Composite Shape Modeling via Latent Space FactorizationAnastasia Dubrovina, Fei Xia, Panos Achlioptas, Mira Shalah et al.ICCV 2019 · 66 citations
- Multi-Class Part Parsing With Joint Boundary-Semantic AwarenessYifan Zhao, Jia Li, Yu Zhang, Yonghong TianICCV 2019 · 63 citations
- PTR: A Benchmark for Part-based Conceptual, Relational, and Physical ReasoningYining Hong, Li Yi, Josh Tenenbaum, Antonio Torralba et al.NeurIPS 2021 · 46 citations
- Hybrid Resolution Network Using Edge Guided Region Mutual Information Loss for Human ParsingYunan Liu, Liang Zhao, Shanshan Zhang, Jian YangACM MM 2020 · 21 citations
Related papers
- Going Denser with Open-Vocabulary Part SegmentationPeize Sun, Shoufa Chen, Chenchen Zhu, Fanyi Xiao et al.ICCV 2023 · 83 citations
- Compositor: Bottom-Up Clustering and Compositing for Robust Part and Object SegmentationJu He, Jieneng Chen, Ming-Xian Lin, Qihang Yu et al.CVPR 2023
- Understanding Multi-Granularity for Open-Vocabulary Part SegmentationJiho Choi, Seonho Lee, Seungho Lee, Minhyun Lee et al.NeurIPS 2024 · 7 citations
- Dynamic Multi-Scale Filters for Semantic SegmentationJunjun He, Zhongying Deng, Yu QiaoICCV 2019 · 287 citations
- PACO: Parts and Attributes of Common ObjectsVignesh Ramanathan, Anmol Kalia, Vladan Petrovic, Yi Wen et al.CVPR 2023
