CustomNet: Object Customization with Variable-Viewpoints in Text-to-Image Diffusion Models
Ziyang Yuan, Mingdeng Cao, Xintao Wang, Zhongang Qi, Chun Yuan, Ying Shan
摘要
Incorporating a customized object into image generation presents an attractive feature in text-to-image (T2I) generation. Some methods finetune T2I models for each object individually at test-time, which tend to be overfitted and time-consuming. Others train an extra encoder to extract object visual information for customization efficiently but struggle to preserve the object's identity. To address these limitations, we present CustomNet, a unified encoder-based object customization framework that explicitly incorporates 3D novel view synthesis capabilities into the customization process. This integration facilitates the adjustment of spatial positions and viewpoints, producing diverse outputs while effectively preserving the object's identity. To train our model effectively, we propose a dataset construction pipeline to better handle real-world objects and complex backgrounds. Additionally, we introduce delicate designs that enable location control and flexible background control through textual descriptions or user-defined backgrounds. Our method allows for object customization without the need of test-time optimization, providing simultaneous control over viewpoints, location, and text. Experimental results show that our method outperforms other customization methods regarding identity preservation, diversity, and harmony. Codes are available at https://github.com/TencentARC/CustomNet.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Generating Multi-Image Synthetic Data for Text-to-Image CustomizationNupur Kumari, Xi Yin, Jun-Yan Zhu, Ishan Misra 等ICCV 2025 · 被引用 1 次
- CustAny: Customizing Anything from A Single ExampleLingjie Kong, Kai Wu, Chengming Xu, Xiaobin Hu 等CVPR 2025
- Preserve Anything: Controllable Image Synthesis with Object PreservationPrasen Kumar Sharma, Neeraj Matiyali, Siddharth Srivastava, Gaurav SharmaICCV 2025
- RealisID: Scale-Robust and Fine-Controllable Identity Customization via Local and Global ComplementationZhaoyang Sun, Fei Du, Weihua Chen, Fan Wang 等AAAI 2025 · 被引用 1 次
- CustomContrast: A Multilevel Contrastive Perspective for Subject-Driven Text-to-Image CustomizationNan Chen, Mengqi Huang, Zhuowei Chen, Yang Zheng 等AAAI 2025 · 被引用 9 次
