MMSite: A Multi-modal Framework for the Identification of Active Sites in Proteins
Song Ouyang, Huiyu Cai, Yong Luo, Kehua Su, Lefei Zhang, Bo Du
摘要
The accurate identification of active sites in proteins is essential for the advancement of life sciences and pharmaceutical development, as these sites are of critical importance for enzyme activity and drug design. Recent advancements in protein language models (PLMs), trained on extensive datasets of amino acid sequences, have significantly improved our understanding of proteins. However, compared to the abundant protein sequence data, functional annotations, especially precise per-residue annotations, are scarce, which limits the performance of PLMs. On the other hand, textual descriptions of proteins, which could be annotated by human experts or a pretrained protein sequence-to-text model, provide meaningful context that could assist in the functional annotations, such as the localization of active sites. This motivates us to construct a ProTein-Attribute text Dataset (ProTAD), comprising over 570,000 pairs of protein sequences and multi-attribute textual descriptions. Based on this dataset, we propose MMSite, a multi-modal framework that improves the performance of PLMs to identify active sites by leveraging biomedical language models (BLMs). In particular, we incorporate manual prompting and design a MACross module to deal with the multi-attribute characteristics of textual descriptions. MMSite is a two-stage ("First Align, Then Fuse") framework: first aligns the textual modality with the sequential modality through soft-label alignment, and then identifies active sites via multi-modal fusion. Experimental results demonstrate that MMSite achieves state-of-the-art performance compared to existing protein representation learning methods. The dataset and code implementation are available at https://github.com/Gift-OYS/MMSite.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Multimodal Protein Language Models for Enzyme Kinetic Parameters: From Substrate Recognition to Conformational AdaptationFei Wang, Xinye Zheng, Kun Li, Yanyan Wei 等CVPR 2026 · 被引用 2 次
- TIGER: Text-Informed Generalized Enzyme-Reaction RetrievalYuhang Zhang, Keyan Ding, Peilin Chen, Han Liu 等ACL 2026
它引用的顶会 Paper15
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 被引用 7,873 次
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 被引用 6,549 次
- Align before Fuse: Vision and Language Representation Learning with Momentum DistillationJunnan Li, Ramprasaath R. Selvaraju, Akhilesh Gotmare, Shafiq R. Joty 等NeurIPS 2021 · 被引用 2,985 次
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch 等ICML 2023 · 被引用 2,601 次
- Language models enable zero-shot prediction of the effects of mutations on protein functionJoshua Meier, Roshan Rao, Robert Verkuil, Jason Liu 等NeurIPS 2021 · 被引用 969 次
相关 Paper
- ProtST: Multi-Modality Learning of Protein Sequences and Biomedical TextsMinghao Xu, Xinyu Yuan, Santiago Miret, Jian TangICML 2023 · 被引用 147 次
- TRIDENT: Tri-Modal Molecular Representation Learning with Taxonomic Annotations and Local CorrespondenceFeng Jiang, Mangal Prakash, Hehuan Ma, Jianyuan Deng 等NeurIPS 2025 · 被引用 5 次
- Prot2Text: Multimodal Protein's Function Generation with GNNs and TransformersHadi Abdine, Michail Chatzianastasis, Costas Bouyioukos, Michalis VazirgiannisAAAI 2024 · 被引用 55 次
- ProtCLIP: Function-Informed Protein Multi-Modal LearningHanjing Zhou, Mingze Yin, Wei Wu, Mingyang Li 等AAAI 2025 · 被引用 11 次
- CrossBind: Collaborative Cross-Modal Identification of Protein Nucleic-Acid-Binding ResiduesLinglin Jing, Sheng Xu, Yifan Wang, Yuzhe Zhou 等AAAI 2024 · 被引用 9 次
