Data Attribution for Text-to-Image Models by Unlearning Synthesized Images
Sheng-Yu Wang, Aaron Hertzmann, Alexei A. Efros, Jun-Yan Zhu, Richard Zhang
摘要
The goal of data attribution for text-to-image models is to identify the training images that most influence the generation of a new image. Influence is defined such that, for a given output, if a model is retrained from scratch without the most influential images, the model would fail to reproduce the same output. Unfortunately, directly searching for these influential images is computationally infeasible, since it would require repeatedly retraining models from scratch. In our work, we propose an efficient data attribution method by simulating unlearning the synthesized image. We achieve this by increasing the training loss on the output image, without catastrophic forgetting of other, unrelated concepts. We then identify training images with significant loss deviations after the unlearning process and label these as influential. We evaluate our method with a computationally intensive but"gold-standard"retraining from scratch and demonstrate our method's advantages over previous methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- EraseFlow: Learning Concept Erasure Policies via GFlowNet-Driven AlignmentNaga Sai Abhiram Kusumba, Maitreya Patel, Kyle Min, Changhoon Kim 等NeurIPS 2025 · 被引用 10 次
- Sharpness-Aware Machine UnlearningHaoran Tang, Rajiv KhannaICLR 2026 · 被引用 10 次
- Fast Data Attribution for Text-to-Image ModelsSheng-Yu Wang, Aaron Hertzmann, Alexei A. Efros, Richard Zhang 等NeurIPS 2025 · 被引用 7 次
- Final-Model-Only Data Attribution with a Unifying View of Gradient-Based MethodsDennis Wei, Inkit Padhi, Soumya Ghosh, Amit Dhurandhar 等NeurIPS 2025 · 被引用 6 次
- Concept-TRAK: Understanding how diffusion models learn concepts through concept attributionYong-Hyun Park, Chieh-Hsin Lai, Satoshi Hayakawa, Yuhta Takida 等ICLR 2026 · 被引用 4 次
它引用的顶会 Paper37
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
相关 Paper
- Training Data Attribution: Was Your Model Secretly Trained On Data Created By Mine?Likun Zhang, Hao Wu, Lingcui Zhang, Fengyuan Xu 等KDD 2025
- Ablating Concepts in Text-to-Image Diffusion ModelsNupur Kumari, Bingliang Zhang, Sheng-Yu Wang, Eli Shechtman 等ICCV 2023 · 被引用 327 次
- Evaluating Data Attribution for Text-to-Image ModelsSheng-Yu Wang, Alexei A. Efros, Jun-Yan Zhu, Richard ZhangICCV 2023 · 被引用 51 次
- Region-Level Data Attribution for Text-To-Image Generative ModelsTrong Bang Nguyen, Phi Le Nguyen, Simon Lucey, Minh HoaiICCV 2025 · 被引用 1 次
- GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via UnlearningNaoki Murata, Yuhta Takida, Chieh-Hsin Lai, Toshimitsu Uesaka 等ICML 2026
