MEDA: Medical-Oriented Activation Editing for Hallucination Mitigation in Medical Large Vision-Language Model
Tianbo Wang, Yuqing Ma, Lingyan Meng, Zhange Zhang, Kewei Liao, Jian Yang, Simin Li, Jinyang Guo, Xianglong Liu
摘要
Despite notable advances in automated medical image interpretation, Medical Large Vision-Language Models (Med-LVLMs) continue to suffer from severe hallucinations, posing critical safety risks in clinical deployment. Editing LVLM activations has shown promise for mitigating hallucination with minimal cost. However, due to the requirements of medical domain expertise, existing methods struggle to capture imaging manifestations and diagnostic principles that are critical for clinical interpretation, thereby limiting their effectiveness. To address these limitations, we propose the first MEDicaloriented Activation Editing (MEDA) method by integrating Query-decisive Manifestation Steering (QMS) and Principle-driven Diagnosis Induction (PDI) to promote Med-LVLM's expertise elicitation. Specifically, QMS retrieves positive query-decisive imaging manifestations as trusted guidance for activation steering, while PDI constructs positive principle-embedded diagnostic prompts to induce expert-like clinical reasoning. Extensive experiments across multiple modalities demonstrate that MEDA efficiently improves the response factuality for both VQA and report generation tasks, achieving up to 10.6% improvement on IU-Xray, while exhibiting generalization and few-shot robustness for practical application.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Inference-Time Intervention: Eliciting Truthful Answers from a Language ModelKenneth Li, Oam Patel, Fernanda B. Viégas, Hanspeter Pfister 等NeurIPS 2023 · 被引用 1,549 次
- LLM-CXR: Instruction-Finetuned LLM for CXR Image Understanding and GenerationSuhyeon Lee, Won Jun Kim, Jinho Chang, Jong Chul YeICLR 2024 · 被引用 80 次
- RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language ModelsPeng Xia, Kangyu Zhu, Haoran Li, Hongtu Zhu 等EMNLP 2024 · 被引用 39 次
- FairCLIP: Harnessing Fairness in Vision-Language LearningYan Luo, Min Shi, Muhammad Osama Khan, Muhammad Muneeb Afzal 等CVPR 2024 · 被引用 37 次
- RaTEScore: A Metric for Radiology Report GenerationWeike Zhao, Chaoyi Wu, Xiaoman Zhang, Ya Zhang 等EMNLP 2024 · 被引用 17 次
相关 Paper
- AFTER: Mitigating the Object Hallucination of LVLM via Adaptive Factual-Guided Activation EditingTianbo Wang, Yuqing Ma, Kewei Liao, Zhange Zhang 等ICLR 2026 · 被引用 2 次
- Query-Routed Activation Editing with Truth-hierarchical Preference OptimizationKewei Liao, Tianbo Wang, Yuqing Ma, Zhange Zhang 等AAAI 2026 · 被引用 1 次
- Beyond Surface Features: Advancing Medical Vision-Language Alignment via Dynamic Evidence-Guided Preference OptimizationZixuan Huang, Zhihong Zhu, Xiaolong Liu, Yanchao Hao 等ACL 2026
- Activation Steering Decoding: Mitigating Hallucination in Large Vision-Language Models through Bidirectional Hidden State InterventionJingran Su, Jingfan Chen, Hongxin Li, Yuntao Chen 等ACL 2025 · 被引用 24 次
- Anatomical Region-Guided Contrastive Decoding: A Plug-and-Play Strategy for Mitigating Hallucinations in Medical VLMsXiao Liang, Chenxi Liu, Zhi Ma, Di Wang 等AAAI 2026 · 被引用 1 次
