Global Human-guided Counterfactual Explanations for Molecular Properties via Reinforcement Learning
Danqing Wang, Antonis Antoniades, Kha-Dinh Luong, Edwin Zhang, Mert Kosan, Jiachen Li, Ambuj K. Singh, William Yang Wang, Lei Li
Abstract
Counterfactual explanations of Graph Neural Networks (GNNs) offer a powerful way to understand data that can naturally be represented by a graph structure. Furthermore, in many domains, it is highly desirable to derive data-driven global explanations or rules that can better explain the high-level properties of the models and data in question. However, evaluating global counterfactual explanations is hard in real-world datasets due to a lack of human-annotated ground truth, which limits their use in areas like molecular sciences. Additionally, the increasing scale of these datasets provides a challenge for random search-based methods. In this paper, we develop a novel global explanation model RLHEX for molecular property prediction. It aligns the counterfactual explanations with human-defined principles, making the explanations more interpretable and easy for experts to evaluate. RLHEX includes a VAE-based graph generator to generate global explanations and an adapter to adjust the latent representation space to human-defined principles. Optimized by Proximal Policy Optimization (PPO), the global explanations produced by RLHEX cover 4.12% more input graphs and reduce the distance between the counterfactual explanation set and the input set by 0.47% on average across three molecular datasets. RLHEX provides a flexible framework to incorporate different human-designed principles into the counterfactual explanation generation process, aligning these explanations with domain expertise. The code and data are released at https://github.com/dqwang122/RLHEX.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2ee3f323-763c-409b-bc88-4768a0e7358dCited by top-tier papers1
Ask how each one uses itBuilds on15
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human SupervisionZhiqing Sun, Yikang Shen, Qinhong Zhou, Hongxin Zhang et al.NeurIPS 2023 · 463 citations
- XGNN: Towards Model-Level Explanations of Graph Neural NetworksHao Yuan, Jiliang Tang, Xia Hu, Shuiwang JiKDD 2020 · 261 citations
- MARS: Markov Molecular Sampling for Multi-objective Drug DiscoveryYutong Xie, Chence Shi, Hao Zhou, Yuwei Yang et al.ICLR 2021 · 186 citations
- Learning and Evaluating Graph Neural Network Explanations based on Counterfactual and Factual ReasoningJuntao Tan, Shijie Geng, Zuohui Fu, Yingqiang Ge et al.WWW 2022 · 151 citations
Related papers
- GnnXemplar: Exemplars to Explanations - Natural Language Rules for Global GNN InterpretabilityBurouj Armgaan, Eshan Jain, Harsh Pandey, Mahesh Chandran et al.NeurIPS 2025 · 5 citations
- DAG Matters! GFlowNets Enhanced Explainer for Graph Neural NetworksWenqian Li, Yinchuan Li, Zhigang Li, Jianye Hao et al.ICLR 2023 · 4 citations
- Reinforcement Learning Enhanced Explainer for Graph Neural NetworksCaihua Shan, Yifei Shen, Yao Zhang, Xiang Li et al.NeurIPS 2021 · 81 citations
- CLEAR: Generative Counterfactual Explanations on GraphsJing Ma, Ruocheng Guo, Saumitra Mishra, Aidong Zhang et al.NeurIPS 2022 · 83 citations
- UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning ModelsHyunju Kang, Geonhee Han, Hogun ParkICLR 2024 · 8 citations
