EditAR: Unified Conditional Generation with Autoregressive Models
Jiteng Mu, Nuno Vasconcelos, Xiaolong Wang
2025Year
9Top-tier citations
Abstract
Change the horse to golden. Change to Cartoon. Remove the bee over the flowers. Change sky to sunset. Change the cat to tiger.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 322e716e-5fae-412c-b6bf-58e4ea2a174bCited by top-tier papers9
- Visual Autoregressive Modeling for Instruction-Guided Image EditingQingyang Mao, Qi Cai, Yehao Li, Yingwei Pan et al.ICLR 2026 · 21 citations
- EditMGT: Unleashing Potentials of Masked Generative Transformers in Image EditingWei Chow, Linfeng Li, Lingdong Kong, Zefeng Li et al.CVPR 2026 · 14 citations
- Semantic Context Matters: Improving Conditioning for Autoregressive ModelsDongyang Jin, Ryan Xu, Jianhao Zeng, Rui Lan et al.CVPR 2026 · 12 citations
- NEP: Autoregressive Image Editing via Next Editing Token PredictionHuimin Wu, Xiaojian (Shawn) Ma, Haozhe Zhao, Yanpeng Zhao et al.NeurIPS 2025 · 8 citations
- OmniGen-AR: AutoRegressive Any-to-Image GenerationJunke Wang, Xun Wang, Qiushan Guo, Peize Sun et al.NeurIPS 2025 · 7 citations
Builds on45
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
Related papers
- SmartEdit: Exploring Complex Instruction-Based Image Editing with Multimodal Large Language ModelsYuzhou Huang, Liangbin Xie, Xintao Wang, Ziyang Yuan et al.CVPR 2024
- UniGen-1.5: Enhancing Image Generation and Editing through Reward Unification in RLRui Tian, Mingfei Gao, Haiming Gang, Jiasen Lu et al.CVPR 2026
- Unveil Inversion and Invariance in Flow Transformer for Versatile Image EditingPengcheng Xu, Boyuan Jiang, Xiaobin Hu, Donghao Luo et al.CVPR 2025
- Align Your Latents: High-Resolution Video Synthesis with Latent Diffusion ModelsAndreas Blattmann, Robin Rombach, Huan Ling, Tim Dockhorn et al.CVPR 2023
- Check, Locate, Rectify: A Training-Free Layout Calibration System for Text- to- Image GenerationBiao Gong, Siteng Huang, Yutong Feng, Shiwei Zhang et al.CVPR 2024
