Knowledge-Guided Object Discovery with Acquired Deep Impressions
Jinyang Yuan, Bin Li, Xiangyang Xue
Abstract
We present a framework called Acquired Deep Impressions (ADI) which continuously learns knowledge of objects as ``impressions'' for compositional scene understanding. In this framework, the model first acquires knowledge from scene images containing a single object in a supervised manner, and then continues to learn from novel multi-object scene images which may contain objects that have not been seen before without any further supervision, under the guidance of the learned knowledge as humans do. By memorizing impressions of objects into parameters of neural networks and applying the generative replay strategy, the learned knowledge can be reused to generate images with pseudo-annotations and in turn assist the learning of novel scenes. The proposed ADI framework focuses on the acquisition and utilization of knowledge, and is complementary to existing deep generative models proposed for compositional scene representation. We adapt a base model to make it fall within the ADI framework and conduct experiments on two types of datasets. Empirical results suggest that the proposed framework is able to effectively utilize the acquired impressions and improve the scene decomposition performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext df64e58a-6449-456a-b82e-afddd0c97e82Cited by top-tier papers2
- Unsupervised Learning of Compositional Scene Representations from Multiple Unspecified ViewpointsJinyang Yuan, Bin Li, Xiangyang XueAAAI 2022 · 12 citations
- Compositional Law Parsing with Latent Random FunctionsFan Shi, Bin Li, Xiangyang XueICLR 2023
Builds on3
- GENESIS: Generative Scene Inference and Sampling with Object-Centric Latent RepresentationsMartin Engelcke, Adam R. Kosiorek, Oiwi Parker Jones, Ingmar PosnerICLR 2020 · 334 citations
- VSGNet: Spatial Attention Network for Detecting Human Object Interactions Using Graph ConvolutionsOytun Ulutan, A. S. M. Iftekhar, B. S. ManjunathCVPR 2020
- Learning to Manipulate Individual Objects in an ImageYanchao Yang, Yutong Chen, Stefano SoattoCVPR 2020
Related papers
- Robust Instance Segmentation Through Reasoning About Multi-Object OcclusionXiaoding Yuan, Adam Kortylewski, Yihong Sun, Alan L. YuilleCVPR 2021
- Continual Learning through Retrieval and ImaginationZhen Wang, Liu Liu, Yiqun Duan, Dacheng TaoAAAI 2022 · 45 citations
- Analogy-Forming Transformers for Few-Shot 3D ParsingNikolaos Gkanatsios, Mayank Singh, Zhaoyuan Fang, Shubham Tulsiani et al.ICLR 2023
- RECALL: Replay-based Continual Learning in Semantic SegmentationAndrea Maracani, Umberto Michieli, Marco Toldo, Pietro ZanuttighICCV 2021 · 148 citations
- Composition-Incremental Learning for Compositional GeneralizationZhen Li, Yuwei Wu, Chenchen Jing, Che Sun et al.AAAI 2026
