TextCraftor: Your Text Encoder can be Image Quality Controller
Yanyu Li, Xian Liu, Anil Kag, Ju Hu, Yerlan Idelbayev, Dhritiman Sagar, Yanzhi Wang, Sergey Tulyakov, Jian Ren
2024年份
12顶会引用
摘要
a cartoon of a house on a mountain a cartoon of a boy playing with a tiger an owl standing on a telephone wire a frustrated child world's best brother t-shirt a girl with long curly blonde hair and sunglasses a bowl with a cartoon dinosaur on it a thumbnail image of a gingerbread man a plate with white rice topped by cooked vegetables Figure 1. Example generated images. For each prompt, we show images generated from three different models, which are SDv1.5, TextCraftor, TextCraftor + UNet, listed from left to right. The random seed is fixed for all generation results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise OptimizationLuca Eyring, Shyamgopal Karthik, Karsten Roth, Alexey Dosovitskiy 等NeurIPS 2024 · 被引用 131 次
- Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion ModelsLuca Eyring, Shyamgopal Karthik, Alexey Dosovitskiy, Nataniel Ruiz 等NeurIPS 2025 · 被引用 36 次
- G-Refine: A General Quality Refiner for Text-to-Image GenerationChunyi Li, Haoning Wu, Hongkun Hao, Zicheng Zhang 等ACM MM 2024 · 被引用 7 次
- Enhancing Motion in Text-to-Video Generation with Decomposed Encoding and ConditioningPenghui Ruan, Pichao Wang, Divya Saxena, Jiannong Cao 等NeurIPS 2024 · 被引用 5 次
- Enhancing Diffusion-based Unrestricted Adversarial Attacks via Adversary Preferences AlignmentKaixun Jiang, Zhaoyu Chen, Haijing Guo, Jinglun Li 等NeurIPS 2025 · 被引用 4 次
它引用的顶会 Paper27
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- StyleDrop: Text-to-Image Synthesis of Any StyleKihyuk Sohn, Lu Jiang, Jarred Barber, Kimin Lee 等NeurIPS 2023 · 被引用 71 次
- Make It Count: Text-to-Image Generation with an Accurate Number of ObjectsLital Binyamin, Yoad Tewel, Hilit Segev, Eran Hirsch 等CVPR 2025
- Latent-NeRF for Shape-Guided Generation of 3D Shapes and TexturesGal Metzer, Elad Richardson, Or Patashnik, Raja Giryes 等CVPR 2023
- Enhancing Compositional Text-to-Image Generation with Reliable Random SeedsShuangqi Li, Hieu Le, Jingyi Xu, Mathieu SalzmannICLR 2025
- ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit AdaptationDar-Yen Chen, Hamish Tennent, Ching-Wen HsuCVPR 2024
