ObjectStitch: Object Compositing with Diffusion Model
Yizhi Song, Zhifei Zhang, Zhe Lin, Scott Cohen, Brian L. Price, Jianming Zhang, Soo Ye Kim, Daniel G. Aliaga
2023Year
69Top-tier citations
Abstract
cal semantics and object appearance. A data augmentation method is further adopted to improve the fidelity of the generator. Our method outperforms relevant baselines in both realism and faithfulness of the synthesized result images in a user study on various real-world images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 10afa662-b428-405f-aba3-641a2e2ea6b6Cited by top-tier papers69
- Zero-shot Image Editing with Reference ImitationXi Chen, Yutong Feng, Mengting Chen, Yiyang Wang et al.NeurIPS 2024 · 80 citations
- GenQuery: Supporting Expressive Visual Search with Generative ModelsKihoon Son, DaEun Choi, Tae Soo Kim, Young-Ho Kim et al.CHI 2024 · 48 citations
- Does FLUX Already Know How to Perform Physically Plausible Image Composition?Shilin Lu, Zhuming Lian, Zihan Zhou, Shaocong Zhang et al.ICLR 2026 · 34 citations
- HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction AwarenessZihui Xue, Romy Luo, Changan Chen, Kristen GraumanNeurIPS 2024 · 29 citations
- Shadows Don't Lie and Lines Can't Bend! Generative Models Don't know Projective Geometry...for NowAyush Sarkar, Hanlin Mai, Amitabh Mahapatra, Svetlana Lazebnik et al.CVPR 2024 · 22 citations
Builds on21
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 6,549 citations
Related papers
- IFCSR: Inference-Free Fidelity-Realism Control for One-Step Diffusion-based Real-World Image Super-ResolutionJonghee Back, Jongju Kim, Jeong-Uk Kim, Eunjin Kim et al.CVPR 2026
- Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified AttentionKyungmin Jo, Jooyeol Yun, Jaegul ChooCVPR 2025
- One Algorithm to Align Them AllBoyi Pang, Savva Ignatyev, Vladimir Ippolitov, Ramil Khafizov et al.CVPR 2026
- Spatially-Invariant Style-Codes Controlled Makeup TransferHan Deng, Chu Han, Hongmin Cai, Guoqiang Han et al.CVPR 2021
- Generative Densification: Learning to Densify Gaussians for High-Fidelity Generalizable 3D ReconstructionSeungtae Nam, Xiangyu Sun, Gyeongjin Kang, Younggeun Lee et al.CVPR 2025
