Contact-Distil: Boosting Low Homologous Protein Contact Map Prediction by Self-Supervised Distillation
Qin Wang, Jiayang Chen, Yuzhe Zhou, Yu Li, Liangzhen Zheng, Sheng Wang, Zhen Li, Shuguang Cui
Abstract
Accurate protein contact map prediction (PCMP) is essential for precise protein structure estimation and further biological studies. Recent works achieve significant performance on this task with high quality multiple sequence alignment (MSA). However, the PCMP accuracy drops dramatically while only poor MSA (e.g., absolute MSA count less than 10) is available. Therefore, in this paper, we propose the Contact-Distil to improve the low homologous PCMP accuracy through knowledge distillation on a self-supervised model. Particularly, two pre-trained transformers are exploited to learn the high quality and low quality MSA representation in parallel for the teacher and student model correspondingly. Besides, the co-evolution information is further extracted from pure sequence through a pretrained ESM-1b model, which provides auxiliary knowledge to improve student performance. Extensive experiments show Contact-Distil outperforms previous state-of-the-arts by large margins on CAMEO-L dataset for low homologous PCMP, i.e., around 13.3% and 9.5% improvements against Alphafold2 and MSA Transformer respectively when MSA count less than 10.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 24cce1ed-0b06-4bfa-8afd-a0f64af3daebCited by top-tier papers1
Ask how each one uses itRelated papers
- PSSM-Distil: Protein Secondary Structure Prediction (PSSP) on Low-Quality PSSM by Knowledge Distillation with Contrastive LearningQin Wang, Boyuan Wang, Zhenlei Xu, Jiaxiang Wu et al.AAAI 2021 · 20 citations
- Co-evolution Transformer for Protein Contact PredictionHe Zhang, Fusong Ju, Jianwei Zhu, Liang He et al.NeurIPS 2021 · 17 citations
- Towards Efficient Pre-Trained Language Model via Feature Correlation DistillationKun Huang, Xin Guo, Meng WangNeurIPS 2023 · 8 citations
- Pixel-Wise Contrastive DistillationJunqiang Huang, Zichao GuoICCV 2023 · 8 citations
- Self-supervised Models are Good Teaching Assistants for Vision TransformersHaiyan Wu, Yuting Gao, Yinqi Zhang, Shaohui Lin et al.ICML 2022 · 26 citations
