HEALNet: Multimodal Fusion for Heterogeneous Biomedical Data
Konstantin Hemker, Nikola Simidjievski, Mateja Jamnik
Abstract
Technological advances in medical data collection, such as high-throughput genomic sequencing and digital high-resolution histopathology, have contributed to the rising requirement for multimodal biomedical modelling, specifically for image, tabular and graph data. Most multimodal deep learning approaches use modality-specific architectures that are often trained separately and cannot capture the crucial cross-modal information that motivates the integration of different data sources. This paper presents the Hybrid Early-fusion Attention Learning Network (HEALNet): a flexible multimodal fusion architecture, which a) preserves modality-specific structural information, b) captures the cross-modal interactions and structural information in a shared latent space, c) can effectively handle missing modalities during training and inference, and d) enables intuitive model inspection by learning on the raw data input instead of opaque embeddings. We conduct multimodal survival analysis on Whole Slide Images and Multi-omic data on four cancer datasets from The Cancer Genome Atlas (TCGA). HEALNet achieves state-of-the-art performance compared to other end-to-end trained fusion models, substantially improving over unimodal and multimodal baselines whilst being robust in scenarios with missing modalities.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8f7ab90d-9ca1-4511-840c-39fbd5fb67baCited by top-tier papers4
- MultiModalPFN: Extending Prior-Data Fitted Networks for Multimodal Tabular LearningWall Kim, Chaeyoung Song, Hanul KimCVPR 2026 · 9 citations
- RAD: Towards Trustworthy Retrieval-Augmented Multi-modal Clinical DiagnosisHaolin Li, Tianjie Dai, Zhe Chen, Siyuan Du et al.NeurIPS 2025 · 3 citations
- Advancing Multimodal Fusion on Heterogeneous Medical Data with Hybrid Geometry AttentionJoy Dhar, Manish Kumar Pandey, Nayyar Zaidi, Chen Chen et al.KDD 2026
- KAMP: Knowledge-Anchored Multimodal Pretraining Framework for Medical Image RepresentationFeiyu Huang, Jia Li, Zhao Chen, Yang Wu et al.CVPR 2026
Builds on4
- Perceiver: General Perception with Iterative AttentionAndrew Jaegle, Felix Gimeno, Andy Brock, Oriol Vinyals et al.ICML 2021 · 1,399 citations
- Multimodal Co-Attention Transformer for Survival Prediction in Gigapixel Whole Slide ImagesRichard J. Chen, Ming Y. Lu, Wei-Hung Weng, Tiffany Y. Chen et al.ICCV 2021 · 369 citations
- Multimodal Optimal Transport-based Co-Attention Transformer with Global Structure Consistency for Survival PredictionYingxue Xu, Hao ChenICCV 2023 · 132 citations
- MultiMoDN - Multimodal, Multi-Task, Interpretable Modular NetworksVinitra Swamy, Malika Satayeva, Jibril Frej, Thierry Bossy et al.NeurIPS 2023 · 27 citations
Related papers
- CA-MLIF: Cross-Attention and Multimodal Low-Rank Interaction Fusion Framework for Tumor Prognostic PredictionYajun An, Jiale Chen, Huan Lin, Zhenbing Liu et al.AAAI 2025 · 1 citation
- Cross-Modal Translation and Alignment for Survival AnalysisFengtao Zhou, Hao ChenICCV 2023 · 123 citations
- Cancer Survival Prediction by Cyclic Generation and Multi-grained AlignmentYongqi Bu, Qinggang Niu, Zhen Li, Yanyu Xu et al.AAAI 2026
- MUST: Modality-Specific Representation-Aware Transformer for Diffusion-Enhanced Survival Prediction with Missing ModalityKyungwon Kim, Dosik HwangCVPR 2026 · 1 citation
- Histopathology-Genomics Multi-modal Structural Representation Learning for Data-Efficient Precision OncologyKun Wu, Zhiguo Jiang, Xinyu Zhu, Jun Shi et al.ICLR 2026
