CrysGNN: Distilling Pre-trained Knowledge to Enhance Property Prediction for Crystalline Materials
Kishalay Das, Bidisha Samanta, Pawan Goyal, Seung-Cheol Lee, Satadeep Bhattacharjee, Niloy Ganguly
Abstract
In recent years, graph neural network (GNN) based approaches have emerged as a powerful technique to encode complex topological structure of crystal materials in an enriched repre- sentation space. These models are often supervised in nature and using the property-specific training data, learn relation- ship between crystal structure and different properties like formation energy, bandgap, bulk modulus, etc. Most of these methods require a huge amount of property-tagged data to train the system which may not be available for different prop- erties. However, there is an availability of a huge amount of crystal data with its chemical composition and structural bonds. To leverage these untapped data, this paper presents CrysGNN, a new pre-trained GNN framework for crystalline materials, which captures both node and graph level structural information of crystal graphs using a huge amount of unla- belled material data. Further, we extract distilled knowledge from CrysGNN and inject into different state of the art prop- erty predictors to enhance their property prediction accuracy. We conduct extensive experiments to show that with distilled knowledge from the pre-trained model, all the SOTA algo- rithms are able to outperform their own vanilla version with good margins. We also observe that the distillation process provides significant improvement over the conventional ap- proach of finetuning the pre-trained model. We will release the pre-trained model along with the large dataset of 800K crys- tal graph which we carefully curated; so that the pre-trained model can be plugged into any existing and upcoming models to enhance their prediction accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 645e80da-3166-472b-ba0e-3dbe848f0b31Cited by top-tier papers6
- A Diffusion-Based Pre-training Framework for Crystal Property PredictionZixing Song, Ziqiao Meng, Irwin KingAAAI 2024 · 25 citations
- LLM Meets Diffusion: A Hybrid Framework for Crystal Material GenerationSubhojyoti Khastagir, Kishalay Das, Pawan Goyal, Seung-Cheol Lee et al.NeurIPS 2025 · 14 citations
- Local-Global Associative Frames for Symmetry-Preserving Crystal Structure ModelingHaowei Hua, Wanyu LinNeurIPS 2025 · 3 citations
- Latent Diffusion Pretraining for Crystal Property PredictionShrimon Mukherjee, KISHALAY DAS, Partha Basuchowdhuri, Pawan Goyal et al.ICML 2026 · 1 citation
- Revisiting the Canonicalization for Fast and Accurate Crystal Tensor Property PredictionHaowei Hua, Jingwen Yang, Wanyu Lin, Pan ZhouAAAI 2026 · 1 citation
Builds on5
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Strategies for Pre-training Graph Neural NetworksWeihua Hu, Bowen Liu, Joseph Gomes, Marinka Zitnik et al.ICLR 2020 · 1,744 citations
- GCC: Graph Contrastive Coding for Graph Neural Network Pre-TrainingJiezhong Qiu, Qibin Chen, Yuxiao Dong, Jing Zhang et al.KDD 2020 · 755 citations
- GPT-GNN: Generative Pre-Training of Graph Neural NetworksZiniu Hu, Yuxiao Dong, Kuansan Wang, Kai-Wei Chang et al.KDD 2020 · 438 citations
- Adversarial Robustness: From Self-Supervised Pre-Training to Fine-TuningTianlong Chen, Sijia Liu, Shiyu Chang, Yu Cheng et al.CVPR 2020
Related papers
- Equivariant Networks for Crystal StructuresSékou-Oumar Kaba, Siamak RavanbakhshNeurIPS 2022 · 38 citations
- Code-Generated Graph Representations Using Multiple LLM Agents for Material Properties PredictionJiao Huang, Qianli Xing, Jinglong Ji, Bo YangICML 2025
- Scalable Multi-Source Pre-training for Graph Neural NetworksMingkai Lin, Wenzhong Li, Xiaobin Hong, Sanglu LuACM MM 2024 · 2 citations
- Extract the Knowledge of Graph Neural Networks and Go Beyond it: An Effective Knowledge Distillation FrameworkCheng Yang, Jiawei Liu, Chuan ShiWWW 2021 · 153 citations
- KPGT: Knowledge-Guided Pre-training of Graph Transformer for Molecular Property PredictionHan Li, Dan Zhao, Jianyang ZengKDD 2022 · 55 citations
