BiGMINT: Biologically-guided Hierarchical Multimodal Integration for Modeling Multiple Compound Activities in Drug Discovery
Pushpak Pati, Bo Li, Abbas Rayabat Khan, Tomé Albuquerque, Steffen Jaensch, Amina Mollaysa, Walid Hassan, Samantha J. Allen, Joke Reumers, Helai P. Mohammad, Scott Oloff, Tommaso Mansi
Abstract
Compound activity modeling is critical for drug discovery, where accurate in silico predictions can significantly reduce reliance on expensive, time‑consuming target-specific experimental assays. Traditional machine learning approaches for compound activity modeling typically rely on either chemoproteomics-centric molecular data or phenotype-centric imaging screens, limiting their ability to capture complementary biological signals. While multimodal approaches show promise, they often fail to capture the interplay between molecular mechanisms and cellular responses. In this paper, we present BiGMINT , a Bi ologically G uided M ultimodal framework that hierarchically INT egrates chemoproteomic and high-content imaging (HCI) data, introducing chemoproteomics-guided phenotypic aggregation, task-aware cross-modal fusion, and protein–protein interaction priors for modeling activities. On two large-scale in-house datasets, with 99K and 40K compound–HCI pairs from U2OS and iNeuron, BiGMINT improves mean AUCROC by up to 10.0% and 4.2%, and high-performing task coverage by up to 103% and 56% over best unimodal and multimodal methods. Thorough analysis revealed mechanistic insights, showing these gains stem from modality complementarity, and protein–protein priors enhance modeling of challenging activities. Code will be released for reproducibility on acceptance of the paper.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a92ea568-fe6c-4e60-a871-2e31ed4e517cBuilds on15
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- FILIP: Fine-grained Interactive Language-Image Pre-TrainingLewei Yao, Runhui Huang, Lu Hou, Guansong Lu et al.ICLR 2022 · 827 citations
- TANKBind: Trigonometry-Aware Neural NetworKs for Drug-Protein Binding Structure PredictionWei Lu, Qifeng Wu, Jixian Zhang, Jiahua Rao et al.NeurIPS 2022 · 254 citations
Related papers
- CellCLIP - Learning Perturbation Effects in Cell Painting via Text-Guided Contrastive LearningMingyu Lu, Ethan Weinberger, Chanwoo Kim, Su-In LeeNeurIPS 2025 · 9 citations
- CP-Agent: Context‑Aware Multimodal Reasoning for Cellular Morphological Profiling under Chemical PerturbationsYuxin Zhang, Yiyao Li, Ping Shu Ho, Simon See et al.ICLR 2026
- FuseMine: Robust Multi-Modal Compound-Protein Interaction Prediction via Differential Attention Feature MiningJunlin Xu, Zhuang Zhang, Zhenghang Gong, Jincan Li et al.AAAI 2026
- KGOT: Unified Knowledge Graph and Optimal Transport Pseudo-Labeling for Molecule-Protein Interaction PredictionJiayu Qin, Zhengquan Luo, Guy Tadmor, Changyou Chen et al.ICLR 2026 · 2 citations
- PETRI: Learning Unified Cell Embeddings from Unpaired Modalities via Early-Fusion Joint ReconstructionRyan W Conrad, Ethan Weinberger, Saradha Venkatachalapathy, Yuwen Chen et al.ICLR 2026
