Segmentation-grounded Scene Graph Generation
Siddhesh Khandelwal, Mohammed Suhail, Leonid Sigal
摘要
Scene graph generation has emerged as an important problem in computer vision. While scene graphs provide a grounded representation of objects, their locations and relations in an image, they do so only at the granularity of proposal bounding boxes. In this work, we propose the first, to our knowledge, framework for pixel-level segmentationgrounded scene graph generation. Our framework is agnostic to the underlying scene graph generation method and address the lack of segmentation annotations in target scene graph datasets (e.g., Visual Genome [23]) through transfer and multi-task learning from, and with, an auxiliary dataset (e.g., MS COCO [28]). Specifically, each target object being detected is endowed with a segmentation mask, which is expressed as a lingual-similarity weighted linear combination over categories that have annotations present in an auxiliary dataset. These inferred masks, along with a novel Gaussian attention mechanism which grounds the relations at a pixel-level within the image, allow for improved relation prediction. The entire framework is end-to-end trainable and is learned in a multi-task manner with both target and auxiliary datasets. * Denotes equal contribution Lady Jacket Bag Car
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- SGTR: End-to-end Scene Graph Generation with TransformerRongjie Li, Songyang Zhang, Xuming HeCVPR 2022 · 被引用 108 次
- Context-aware Scene Graph Generation with Seq2Seq TransformersYichao Lu, Himanshu Rai, Jason Chang, Boris Knyazev 等ICCV 2021 · 被引用 93 次
- Iterative Scene Graph GenerationSiddhesh Khandelwal, Leonid SigalNeurIPS 2022 · 被引用 47 次
- HiLo: Exploiting High Low Frequency Relations for Unbiased Panoptic Scene Graph GenerationZijian Zhou, Miaojing Shi, Holger CaesarICCV 2023 · 被引用 29 次
- TextPSG: Panoptic Scene Graph Generation from Textual DescriptionsChengyang Zhao, Yikang Shen, Zhenfang Chen, Mingyu Ding 等ICCV 2023 · 被引用 24 次
它引用的顶会 Paper7
- Unpaired Image Captioning via Scene Graph AlignmentsJiuxiang Gu, Shafiq R. Joty, Jianfei Cai, Handong Zhao 等ICCV 2019 · 被引用 191 次
- Uncertainty-Aware Learning for Zero-Shot Semantic SegmentationPing Hu, Stan Sclaroff, Kate SaenkoNeurIPS 2020 · 被引用 79 次
- Action Genome: Actions As Compositions of Spatio-Temporal Scene GraphsJingwei Ji, Ranjay Krishna, Li Fei-Fei, Juan Carlos NieblesCVPR 2020
- GPS-Net: Graph Property Sensing Network for Scene Graph GenerationXin Lin, Changxing Ding, Jinquan Zeng, Dacheng TaoCVPR 2020
- UniT: Unified Knowledge Transfer for Any-Shot Object Detection and SegmentationSiddhesh Khandelwal, Raghav Goyal, Leonid SigalCVPR 2021
相关 Paper
- Part-Aware Interactive Learning for Scene Graph GenerationHongshuo Tian, Ning Xu, An-An Liu, Yongdong ZhangACM MM 2020 · 被引用 12 次
- Scene Graph Prediction With Limited LabelsRanjay Krishna, Vincent S. Chen, Paroma Varma, Michael S. Bernstein 等ICCV 2019 · 被引用 5 次
- Focusing on Flexible Masks: A Novel Framework for Panoptic Scene Graph Generation with Relation ConstraintsJiarui Yang, Chuan Wang, Zeming Liu, Jiahong Wu 等ACM MM 2023 · 被引用 8 次
- Weakly-supervised Video Scene Graph Generation via Unbiased Cross-modal LearningZiyue Wu, Junyu Gao, Changsheng XuACM MM 2023 · 被引用 5 次
- EGTR: Extracting Graph from Transformer for Scene Graph GenerationJinbae Im, JeongYeon Nam, Nokyung Park, Hyungmin Lee 等CVPR 2024
