Bridging the Modality Reliability Gap in Drug-Target Interaction Prediction via a Confidence-aware Multimodal Fusion Framework
Jie Yang, Junxiong Zhang, Kun Qian, Qingyu Yang, Weikai Li, Zhen Cheng
Abstract
With the rapid advancement of deep learning, drug target interaction (DTI) prediction has seen substantial performance enhancements. However, existing methodologies face a critical, yet unaddressed challenge, i.e., the Modality Reliability Gap. Such a gap arises from the unpredictable variance in the informativeness and reliability of 1D sequence versus 3D structural data across different drug-target pairs, critically limiting model robustness and domain generalization capabilities. To overcome it, we introduce DrugCMF, a novel Drug-Target interaction prediction method via Confidence-aware Multimodal Fusion framework designed specifically to bridge the Modality Reliability Gap. Specifically, the DrugCMF employs a four-stage approach: (1) it extracts rich features by utilizing four pre-trained models to obtain token-level embeddings from both 1D sequences and 3D structures. (2) it preserves modality informativeness by independently learning interaction patterns within each modality through a Token-level Interaction module. (3) it explicitly quantifies the reliability gap by employing a novel confidence estimation mechanism to dynamically learn weights for each modality. (4) it bridges the gap by using these confidence scores to guide a learnable cross-modal fusion module, adaptively fusing information from the most trustworthy source. By methodically addressing the Modality Reliability Gap, DrugCMF significantly outperforms SOTA methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 92a5e793-9119-4f49-a86d-5308541dc368Builds on19
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 2,213 citations
- Perceiver: General Perception with Iterative AttentionAndrew Jaegle, Felix Gimeno, Andy Brock, Oriol Vinyals et al.ICML 2021 · 1,399 citations
- SaProt: Protein Language Modeling with Structure-aware VocabularyJin Su, Chenchen Han, Yuyang Zhou, Junjie Shan et al.ICLR 2024 · 285 citations
- Uni-Mol: A Universal 3D Molecular Representation Learning FrameworkGengmo Zhou, Zhifeng Gao, Qiankun Ding, Hang Zheng et al.ICLR 2023 · 254 citations
- Training independent subnetworks for robust predictionMarton Havasi, Rodolphe Jenatton, Stanislav Fort, Jeremiah Zhe Liu et al.ICLR 2021 · 235 citations
Related papers
- R-DTI: Drug Target Interaction Prediction Based on Second-Order Relevance ExplorationYang Hua, Tianyang Xu, Xiaoning Song, Zhenhua Feng et al.AAAI 2025 · 2 citations
- FuseMine: Robust Multi-Modal Compound-Protein Interaction Prediction via Differential Attention Feature MiningJunlin Xu, Zhuang Zhang, Zhenghang Gong, Jincan Li et al.AAAI 2026
- GRAM-DTI: Adaptive Multimodal Representation Learning for Drug–Target Interaction PredictionFeng Jiang, Amina Mollaysa, Hehuan Ma, Yuzhi Guo et al.ICLR 2026
- CMF-ELN: A Cross-Modal-Fused End-to-end Learning Network for Cold-Start Drug-Drug Interaction PredictionDi Wu, Hongyi Sun, Haichao Xu, Jia Chen et al.KDD 2026
- DPNET: Dynamic Poly-attention Network for Trustworthy Multi-modal ClassificationXin Zou, Chang Tang, Xiao Zheng, Zhenglai Li et al.ACM MM 2023 · 16 citations
