PIRN: Prototypical-based Intra-modal Reconstruction with Normality Communication for Multi-modal Anomaly Detection.
YITING LI, Xulei Yang, Jing Zhang, Sichao Tian, Jingyi Liao, Fayao Liu
摘要
Unsupervised multimodal anomaly detection (MAD) aims to detect anomalies by using both RGB and 3D modalities. However, existing methods struggle in few-shot scenarios where the number of normal training samples is limited. Specifically, cross-modal alignment approaches fail to learn reliable correspondences from scarce normal data, whereas memory-based methods often misclassify unseen normal variations as anomalies. To address these issues, we propose , a prototype-driven reconstruction framework equipped with explicit cross-modal knowledge transfer. Instead of relying on dense feature alignment or heavy memory banks, uses a compact set of learnable prototypes to capture diverse normal patterns and constrain feature reconstruction. Specifically, our framework incorporates three core innovations. We introduce Balanced Prototype Assignment (BPA), which employs balanced optimal transport to ensure uniform prototype utilization and prevent codebook collapse. Next, we propose Adaptive Prototype Refinement (APR), which uses gated prototype updates to dynamically expand the model's knowledge of unseen normal variations during inference. To enable each modality to assist the other in reconstructing, we further develop a Multimodal Normality Communication (MNC) module that exchanges high-level normal cues between modalities via gated cross-attention. Extensive experiments on the MVTec 3D-AD, Eyecandies, and Real-IAD benchmarks validate the effectiveness of , where it consistently achieves superior performance compared to existing baselines under challenging few-shot settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper21
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Hierarchical Vector Quantized Transformer for Multi-class Unsupervised Anomaly DetectionRuiying Lu, Yujie Wu, Long Tian, Dongsheng Wang 等NeurIPS 2023 · 被引用 121 次
- FastRecon: Few-shot Industrial Anomaly Detection via Fast Feature ReconstructionZheng Fang, Xiaoyang Wang, Haocheng Li, Jiejie Liu 等ICCV 2023 · 被引用 100 次
- Shape-Guided Dual-Memory Learning for 3D Anomaly DetectionYu-Min Chu, Chieh Liu, Ting-I Hsieh, Hwann-Tzong Chen 等ICML 2023 · 被引用 80 次
- Online Prototype Learning for Online Continual LearningYujie Wei, Jiaxin Ye, Zhizhong Huang, Junping Zhang 等ICCV 2023 · 被引用 78 次
相关 Paper
- Remove the Ambiguity: Few-shot Multimodal Anomaly Detection Using Crossmodal Feature ReplacersYuan Guo, Wanqi Zhang, Xu WangICML 2026
- FastRef: Fast Prototype Refinement for Few-shot Industrial Anomaly DetectionYufei Li, Long Tian, Yuyang Dai, Wenchao Chen 等CVPR 2026 · 被引用 7 次
- Is Task-Specific Training Necessary for Anomaly Detection?Xingwu Zhang, Guanxuan Li, Paul Henderson, Gerardo Aragon-Camarasa 等ICML 2026 · 被引用 1 次
- Complementary Prototype Mapping for Efficient Multimodal Anomaly DetectionYuan Zhao, Zhang xiaoqin to Xiaoqin Zhang, Huchuan Lu, Lihe ZhangCVPR 2026
- Beyond Single-Modal Boundary: Cross-Modal Anomaly Detection through Visual Prototype and HarmonizationKai Mao, Ping Wei, Yiyang Lian, Yangyang Wang 等CVPR 2025
