Edicho: Consistent Image Editing in the Wild
Qingyan Bai, Hao Ouyang, Yinghao Xu, Qiuyu Wang, Ceyuan Yang, Ka Leong Cheng, Yujun Shen, Qifeng Chen
Abstract
As a verified need, consistent editing across in-the-wild images remains a technical challenge arising from various unmanageable factors, like object poses, lighting conditions, and photography environments. Edicho11“Edicho” is an abbreviation of “edit echo”, implying that the edit is echoed across images. steps in with a training-free solution based on diffusion models, featuring a fundamental design principle of using explicit image correspondence to direct editing. Specifically, the key components include an attention manipulation module and a carefully refined classifier-free guidance () denoising strategy, both of which take into account the preestimated correspondence. Such an inference-time algorithm enjoys a plug-and-play nature and is compatible to most diffusion-based editing methods, such as ControlNet and BrushNet. Extensive results demonstrate the efficacy of Edicho in consistent cross-image editing under diverse settings. Project page can be found here.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 27c341de-766b-4752-85cd-8b85bbda9010Cited by top-tier papers4
- Group Editing: Edit Multiple Images in One GoYue Ma, Xinyu Wang, Qianli Ma, Qinghe Wang et al.CVPR 2026 · 15 citations
- Mind-the-Glitch: Visual Correspondence for Detecting Inconsistencies in Subject-Driven GenerationAbdelrahman Eldesokey, Aleksandar Cvejic, Bernard Ghanem, Peter WonkaNeurIPS 2025 · 6 citations
- RewardFlow: Generate Images by Optimizing What You RewardOnkar Susladkar, Dong-Hwan Jang, Tushar Prakash, Adheesh Sunil Juvekar et al.CVPR 2026 · 2 citations
- Match-and-Fuse: Consistent Generation from Unstructured Image SetsKate Feingold, Omri Kaduri, Tali DekelCVPR 2026
Builds on45
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and GenerationLiyao Jiang, Negar Hassanpour, Mohammad Salameh, Mohammadreza Samadi et al.AAAI 2025 · 8 citations
- TweezeEdit: Consistent and Efficient Image Editing with Path RegularizationJianda Mao, Kaibo Wang, Yang Xiang, Kani ChenAAAI 2026 · 2 citations
- Inversion-Free Image Editing with Language-Guided Diffusion ModelsSihan Xu, Yidong Huang, Jiayi Pan, Ziqiao Ma et al.CVPR 2024 · 12 citations
- PostEdit: Posterior Sampling for Efficient Zero-Shot Image EditingFeng Tian, Yixuan Li, Yichao Yan, Shanyan Guan et al.ICLR 2025
- Training-Free Text-Guided Color Editing with Multi-Modal Diffusion TransformerZixin Yin, Xili Dai, Ling-Hao Chen, Deyu Zhou et al.ICLR 2026 · 6 citations
