SeaS: Few-Shot Industrial Anomaly Image Generation with Separation and Sharing Fine-Tuning
Zhewei Dai, Shilei Zeng, Haotian Liu, Xurui Li, Feng Xue, Yu Zhou
摘要
We introduce SeaS, a unified industrial generative model for automatically creating diverse anomalies, authentic normal products, and precise anomaly masks. While extensive research exists, most efforts either focus on specific tasks, i.e., anomalies or normal products only, or require separate models for each anomaly type. Consequently, prior methods either offer limited generative capability or depend on a vast array of anomaly-specific models. We demonstrate that U-Net's differentiated learning ability captures the distinct visual traits of slightly-varied normal products and diverse anomalies, enabling us to construct a unified model for all tasks. Specifically, we first introduce an Unbalanced Abnormal (UA) Text Prompt, comprising one normal token and multiple anomaly tokens. More importantly, our Decoupled Anomaly Alignment (DA) loss decouples anomaly attributes and binds them to distinct anomaly tokens of UA, enabling SeaS to create unseen anomalies by recombining these attributes. Furthermore, our Normal-image Alignment (NA) loss aligns the normal token to normal patterns, making generated normal products globally consistent and locally varied. Finally, SeaS produces accurate anomaly masks by fusing discriminative U-Net features with high-resolution VAE features. SeaS sets a new benchmark for industrial generation, significantly enhancing downstream applications, with average improvements of pixel-level AP for synthesis-based AD approaches, image-level AP for unsupervised AD methods, and IoU for supervised segmentation models. Code is available at https://github.com/HUST-SLOW/SeaS.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Towards an Incremental Unified Multimodal Anomaly Detection: Augmenting Multimodal Denoising From an Information Bottleneck PerspectiveKaifang Long, Lianbo Ma, Jiaqi Liu, liming liu 等CVPR 2026 · 被引用 5 次
- One-to-More: High-Fidelity Training-Free Anomaly Generation with Attention ControlHaoxiang Rao, Zhao Wang, Chenyang Si, Yan Lyu 等CVPR 2026 · 被引用 2 次
- RAID: Retrieval-Augmented Anomaly DetectionMingxiu Cai, Zhe Zhang, Gaochang Wu, Tianyou Chai 等CVPR 2026 · 被引用 2 次
- Anomaly-Preference Image GenerationFuyun Wang, Yuanzhi Wang, Xu Guo, Sujia Huang 等ICML 2026 · 被引用 1 次
- CHIMERA: Controllable High-quality Image-Mask Extraction for Reliable Diffusion-based Anomaly SynthesisJoungBin Lee, Hyunkoo Lee, Jini Yang, Chaehyun Kim 等AAAI 2026
它引用的顶会 Paper23
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Towards Total Recall in Industrial Anomaly DetectionKarsten Roth, Latha Pemula, Joaquin Zepeda, Bernhard Schölkopf 等CVPR 2022 · 被引用 1,301 次
- Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion ModelsHila Chefer, Yuval Alaluf, Yael Vinker, Lior Wolf 等SIGGRAPH 2023 · 被引用 438 次
- SVDiff: Compact Parameter Space for Diffusion Fine-TuningLigong Han, Yinxiao Li, Han Zhang, Peyman Milanfar 等ICCV 2023 · 被引用 384 次
相关 Paper
- AnomalyControl: Highly-Aligned Anomalous Image Generation with Controlled Diffusion ModelYuanyi Duan, Wei Xu, Qinlong Wu, Guo-Sen Xie 等ACM MM 2025 · 被引用 3 次
- Generate Aligned Anomaly: Region-Guided Few-Shot Anomaly Image-Mask Pair Synthesis for Industrial InspectionYilin Lu, Jianghang Lin, Linhuang Xie, Kai Zhao 等ACM MM 2025
- UniAd: Unified Adversarial Alignment for Unsupervised Cross-Domain Industrial Anomaly DetectionYulong Fang, Zhanshan Li, Jingyao LiKDD 2026
- FAST: Foreground-aware Diffusion with Accelerated Sampling Trajectory for Segmentation-oriented Anomaly SynthesisXichen Xu, Yanshu Wang, Jinbao Wang, Xiaoning Lei 等NeurIPS 2025 · 被引用 1 次
- OmniAL: A Unified CNN Framework for Unsupervised Anomaly LocalizationYing ZhaoCVPR 2023
