BiGMINT: Biologically-guided Hierarchical Multimodal Integration for Modeling Multiple Compound Activities in Drug Discovery
Pushpak Pati, Bo Li, Abbas Rayabat Khan, Tomé Albuquerque, Steffen Jaensch, Amina Mollaysa, Walid Hassan, Samantha J. Allen, Joke Reumers, Helai P. Mohammad, Scott Oloff, Tommaso Mansi
摘要
Compound activity modeling is critical for drug discovery, where accurate in silico predictions can significantly reduce reliance on expensive, time‑consuming target-specific experimental assays. Traditional machine learning approaches for compound activity modeling typically rely on either chemoproteomics-centric molecular data or phenotype-centric imaging screens, limiting their ability to capture complementary biological signals. While multimodal approaches show promise, they often fail to capture the interplay between molecular mechanisms and cellular responses. In this paper, we present BiGMINT , a Bi ologically G uided M ultimodal framework that hierarchically INT egrates chemoproteomic and high-content imaging (HCI) data, introducing chemoproteomics-guided phenotypic aggregation, task-aware cross-modal fusion, and protein–protein interaction priors for modeling activities. On two large-scale in-house datasets, with 99K and 40K compound–HCI pairs from U2OS and iNeuron, BiGMINT improves mean AUCROC by up to 10.0% and 4.2%, and high-performing task coverage by up to 103% and 56% over best unimodal and multimodal methods. Thorough analysis revealed mechanistic insights, showing these gains stem from modality complementarity, and protein–protein priors enhance modeling of challenging activities. Code will be released for reproducibility on acceptance of the paper.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper15
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- FILIP: Fine-grained Interactive Language-Image Pre-TrainingLewei Yao, Runhui Huang, Lu Hou, Guansong Lu 等ICLR 2022 · 被引用 827 次
- TANKBind: Trigonometry-Aware Neural NetworKs for Drug-Protein Binding Structure PredictionWei Lu, Qifeng Wu, Jixian Zhang, Jiahua Rao 等NeurIPS 2022 · 被引用 254 次
相关 Paper
- CellCLIP - Learning Perturbation Effects in Cell Painting via Text-Guided Contrastive LearningMingyu Lu, Ethan Weinberger, Chanwoo Kim, Su-In LeeNeurIPS 2025 · 被引用 9 次
- CP-Agent: Context‑Aware Multimodal Reasoning for Cellular Morphological Profiling under Chemical PerturbationsYuxin Zhang, Yiyao Li, Ping Shu Ho, Simon See 等ICLR 2026
- FuseMine: Robust Multi-Modal Compound-Protein Interaction Prediction via Differential Attention Feature MiningJunlin Xu, Zhuang Zhang, Zhenghang Gong, Jincan Li 等AAAI 2026
- KGOT: Unified Knowledge Graph and Optimal Transport Pseudo-Labeling for Molecule-Protein Interaction PredictionJiayu Qin, Zhengquan Luo, Guy Tadmor, Changyou Chen 等ICLR 2026 · 被引用 2 次
- PETRI: Learning Unified Cell Embeddings from Unpaired Modalities via Early-Fusion Joint ReconstructionRyan W Conrad, Ethan Weinberger, Saradha Venkatachalapathy, Yuwen Chen 等ICLR 2026
