MoEdit: On Learning Quantity Perception for Multi-object Image Editing
Yanfeng Li, Ka-Hou Chan, Yue Sun, Chan-Tong Lam, Tong Tong, Zitong Yu, Keren Fu, Xiaohong Liu, Tao Tan
2025年份
摘要
Ten koalas "...steampunck style" "...vibrant portrait painting of Salvador Dalí" "...with blanket" "→mice, by the sea" "→foxes, futuristic metropolis style" Reference TurboEdit MoEdit (Ours) Three rabbits and two foxes "...dark horror style" "...with smilling faces" "→bears, in a natural field" "→corgis, by the sea" "→foxes, in a natural field" Figure 1. Visual comparisons of our MoEdit with TurboEdit [52]. Reference represents input images. Five different images edited by each method are based on five distinct text prompts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper36
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann 等ICLR 2024 · 被引用 4,569 次
- CLIPScore: A Reference-free Evaluation Metric for Image CaptioningJack Hessel, Ari Holtzman, Maxwell Forbes, Ronan Le Bras 等EMNLP 2021 · 被引用 937 次
- BLIP-Diffusion: Pre-trained Subject Representation for Controllable Text-to-Image Generation and EditingDongxu Li, Junnan Li, Steven C. H. HoiNeurIPS 2023 · 被引用 587 次
- Uni-ControlNet: All-in-One Control to Text-to-Image Diffusion ModelsShihao Zhao, Dongdong Chen, Yen-Chun Chen, Jianmin Bao 等NeurIPS 2023 · 被引用 505 次
相关 Paper
- CCEdit: Creative and Controllable Video Editing via Diffusion ModelsRuoyu Feng, Wenming Weng, Yanhui Wang, Yuhui Yuan 等CVPR 2024
- HQ-Edit: A High-Quality Dataset for Instruction-based Image EditingMude Hui, Siwei Yang, Bingchen Zhao, Yichun Shi 等ICLR 2025
- DreamCatalyst: Fast and High-Quality 3D Editing via Controlling Editability and Identity PreservationJiwook Kim, Seonho Lee, Jaeyo Shin, Jiho Choi 等ICLR 2025
- RAVE: Randomized Noise Shuffling for Fast and Consistent Video Editing with Diffusion ModelsOzgur Kara, Bariscan Kurtkaya, Hidir Yesiltepe, James M. Rehg 等CVPR 2024
- Style-Editor: Text-driven Object-centric Style EditingJihun Park, Jongmin Gim, Kyoungmin Lee, Seunghun Lee 等CVPR 2025
