Exploring Sparse MoE in GANs for Text-conditioned Image Synthesis
Jiapeng Zhu, Ceyuan Yang, Kecheng Zheng, Yinghao Xu, Zifan Shi, Yifei Zhang, Qifeng Chen, Yujun Shen
Abstract
A black and white photo of a dog with big eyes. A steaming plate of classic fried rice with vegetables, eggs. A photo of a colorful sports car, with mountain background. Oil painting of a flower garden, distant cottage, soft sunlight. Sunrise over the clouds in the mountains. A pencil sketch of a woman with deep, almond-shaped eyes in. A coral reef with starfish, sea anemones and fish. Figure 1. Example results at 512×512 resolution synthesized by our proposed Aurora, a large-scale GAN-based text-to-image generator.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext aa6623f9-4597-4c6d-b2b8-4308bbade393Cited by top-tier papers5
- Adversarial Flow ModelsShanchuan Lin, Ceyuan Yang, Zhijie Lin, Hao Chen et al.ICML 2026 · 8 citations
- Towards Scalable Topological RegularizersHiu-Tung Wong, Darrick Lee, Hong YanICLR 2025
- X-Fusion: Introducing New Modality to Frozen Large Language ModelsSicheng Mo, Thao Nguyen, Xun Huang, Siddharth Srinivasan Iyer et al.ICCV 2025
- Scalable GANs with TransformersSangeek Hyun, MinKyu Lee, Jae-Pil HeoICML 2026
- Test-time Domain Generalization for Image Super-resolutionZaizuo Tang, Yu-Bin YangICLR 2026
Builds on49
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- UniGen-1.5: Enhancing Image Generation and Editing through Reward Unification in RLRui Tian, Mingfei Gao, Haiming Gang, Jiasen Lu et al.CVPR 2026
- TextCraftor: Your Text Encoder can be Image Quality ControllerYanyu Li, Xian Liu, Anil Kag, Ju Hu et al.CVPR 2024
- Multi-Concept Customization of Text-to-Image DiffusionNupur Kumari, Bingliang Zhang, Richard Zhang, Eli Shechtman et al.CVPR 2023
- Chat2SVG: Vector Graphics Generation with Large Language Models and Image Diffusion ModelsRonghuan Wu, Wanchao Su, Jing LiaoCVPR 2025
- Align Your Latents: High-Resolution Video Synthesis with Latent Diffusion ModelsAndreas Blattmann, Robin Rombach, Huan Ling, Tim Dockhorn et al.CVPR 2023
