E³SAM2: Entropy-Aware and Edge-Guided Adaptation of SAM2 for Echocardiography Video Segmentation
Long Zheng, Zhi Li, Weidong Wang, Zhenyu Dai, Shuyun Li
Abstract
Foundation segmentation models, such as SAM and its video-oriented variant SAM2, have achieved remarkable success in natural image and video segmentation. However, their direct application to echocardiography video is challenged by structural uncertainty arising from severe speckle noise and blurry anatomical boundaries. To address this, we propose E³SAM2, a lightweight adaptation framework that introduces a novel entropy-based methodology to explicitly model and mitigate such uncertainty. Specifically, an entropy-guided attention mechanism is introduced to steer the model’s focus toward structurally reliable features, particularly in speckle-dominated regions. Additionally, an entropy regularization loss is introduced to further enhance target-background discrimination. To better resolve indistinct anatomical contours, an edge-aware supervision module is incorporated to inject explicit boundary priors for sharper delineation. These components are efficiently integrated through a global-local feature adapter. Experiments on CAMUS and EchoNet-Dynamic datasets demonstrate that E³SAM2 achieves state-of-the-art segmentation and clinical estimation performance, while maintaining high computational efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6dfa9f8f-73cf-46c9-8118-d5d5c9154941Builds on3
- Multi-Scale and Detail-Enhanced Segment Anything Model for Salient Object DetectionShixuan Gao, Pingping Zhang, Tianyu Yan, Huchuan LuACM MM 2024 · 93 citations
- Super-efficient Echocardiography Video Segmentation via Proxy- and Kernel-Based Semi-supervised LearningHuisi Wu, Jingyin Lin, Wende Xie, Jing QinAAAI 2023 · 16 citations
- Learning A Sparse Transformer Network for Effective Image DerainingXiang Chen, Hao Li, Mingqiang Li, Jinshan PanCVPR 2023
Related papers
- MemSAM: Taming Segment Anything Model for Echocardiography Video SegmentationXiaolong Deng, Huisi Wu, Runhao Zeng, Jing QinCVPR 2024
- Semi-supervised Echocardiography Video Segmentation via Anchor Semantic Awareness and Continuous Pseudo-label ReforgingYunpeng Fang, Yimu Sun, Jingxing Guo, Huisi Wu et al.CVPR 2026
- Hierarchical Spatiotemporal Context Aggregation and Speckle-aware Deformable Convolution for Echocardiography Video SegmentationJingxing Guo, Guilian Chen, Yimu Sun, Huisi Wu et al.ACM MM 2025
- EchoVim: Making Vision Mamba Docile for Echocardiography Video Segmentation via Dynamic Interaction and Semantic Token-attentive RefinementJingxing Guo, Guilian Chen, Yimu Sun, Huisi Wu et al.ACM MM 2025
- Semi-supervised TEE Segmentation via Interacting with SAM Equipped with Noise-Resilient PromptingSen Deng, Yidan Feng, Haoneng Lin, Yiting Fan et al.AAAI 2024 · 3 citations
