MeronymNet: A Hierarchical Model for Unified and Controllable Multi-Category Object Generation
Rishabh Baghel, Abhishek Trivedi, Tejas Ravichandran, Ravi Kiran Sarvadevabhatla
Abstract
We introduce MeronymNet, a novel hierarchical approach for controllable, part-based generation of multi-category objects using a single unified model. We adopt a guided coarse-to-fine strategy involving semantically conditioned generation of bounding box layouts, pixel-level part layouts and ultimately, the object depictions themselves. We use Graph Convolutional Networks, Deep Recurrent Networks along with custom-designed Conditional Variational Autoencoders to enable flexible, diverse and category-aware generation of 2-D objects in a controlled manner. The performance scores for generated objects reflect MeronymNet's superior performance compared to multiple strong baselines and ablative variants. We also showcase MeronymNet's suitability for controllable object generation and interactive object editing at various levels of structural and semantic granularity.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e4b8f852-63f2-4c0f-9ad8-6763590ae294Builds on3
- LayoutVAE: Stochastic Scene Layout Generation From a Label SetAkash Abdu Jyothi, Thibaut Durand, Jiawei He, Leonid Sigal et al.ICCV 2019 · 194 citations
- Image Synthesis From Reconfigurable Layout and StyleWei Sun, Tianfu WuICCV 2019 · 160 citations
- PQ-NET: A Generative Part Seq2Seq Network for 3D ShapesRundi Wu, Yixin Zhuang, Kai Xu, Hao Zhang et al.CVPR 2020
Related papers
- PLATO: Generating Objects from Part Lists via Synthesized LayoutsAmruta Muthal, Varghese P. Kuruvilla, Ravi Kiran SarvadevabhatlaACM MM 2025
- DC-ControlNet: Decoupling Inter- and Intra-Element Conditions in Image Generation with Diffusion ModelsHongji Yang, Wencheng Han, Yucheng Zhou, Jianbing ShenICCV 2025 · 4 citations
- Learning Part Generation and Assembly for Structure-Aware Shape SynthesisJun Li, Chengjie Niu, Kai XuAAAI 2020 · 85 citations
- StructEdit: Learning Structural Shape VariationsKaichun Mo, Paul Guerrero, Li Yi, Hao Su et al.CVPR 2020
- LayoutTransformer: Scene Layout Generation With Conceptual and Spatial DiversityCheng-Fu Yang, Wan-Cyuan Fan, Fu-En Yang, Yu-Chiang Frank WangCVPR 2021
