DISSOLVR: An Interpretable and Fast Framework for Aqueous and Organic Solubility Prediction
Vansh Ramani, Har A Arora, Dhairya Kuchhal, Sayan Ranu, Tarak Karmakar
Abstract
High-fidelity solubility prediction is fundamental to pharmaceutical development and environmental partitioning, where accurate modeling must couple molecular structure with thermodynamic behavior across diverse chemical environments. However, recent advancements have been dominated by deep learning architectures that often sacrifice physical interpretability for predictive power. We challenge this trend by showing that state-of-the-art performance does not require such non-transparent architectures. To address this, we introduce DISSOLVR, a transparent framework for molecular solubility prediction. In addition, we perform a comprehensive literature review and a benchmarking study against various methods. We show that DISSOLVR approaches the aleatoric limit of experimental uncertainty and achieves OOD generalization through structural invariance, derived by mapping molecules to physically-grounded descriptors. Then, we present an LLM-assisted post-hoc explanation pipeline that bridges the gap between symbolic model artifacts and chemically grounded narratives. Finally, a comparative benchmark of a survey involving 22 expert chemists reveals that expert evaluators provide deep insights.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on2
Related papers
- Chemically Interpretable Graph Interaction Network for Prediction of Pharmacokinetic Properties of Drug-Like MoleculesYashaswi Pathak, Siddhartha Laghuvarapu, Sarvesh Mehta, U. Deva PriyakumarAAAI 2020 · 45 citations
- Guiding Deep Molecular Optimization with Genetic ExplorationSungsoo Ahn, Junsu Kim, Hankook Lee, Jinwoo ShinNeurIPS 2020 · 98 citations
- MolecularIQ: Characterizing Chemical Reasoning Capabilities Through Symbolic Verification on Molecular GraphsChristoph Bartmann, Johannes Schimunek, Mykyta Ielanskyi, Philipp Seidl et al.ICLR 2026 · 5 citations
- ConSim: Measuring Concept-Based Explanations' Effectiveness with Automated SimulatabilityAntonin Poché, Alon Jacovi, Agustin Martin Picard, Victor Boutin et al.ACL 2025 · 8 citations
- MolRAG: Unlocking the Power of Large Language Models for Molecular Property PredictionZiting Xian, Jiawei Gu, Lingbo Li, Shangsong LiangACL 2025 · 7 citations
