Credible Information Subset Decomposition: An End-to-End Multi-fidelity Learning Model by Modeling Label Information
Sihan Wang, Wenjie Du, Yang Wang
Abstract
In the AI4Chemistry scenario, utilizing heterogeneous data at different fidelity levels is a common and core issue. High-fidelity data is accurate but scarce, while low-fidelity data is abundant but biased. Traditional multi-fidelity methods typically identify cross-fidelity biases based on paired samples under different fidelity labels. However, due to the mismatch in dataset input distribution and the complexity of the biases themselves, these methods are difficult to implement in real-world scientific environments. To address this, we propose a trusted information subset decomposition framework that can efficiently utilize multi-fidelity data without requiring paired samples. Multi-fidelity label supervision is decomposed into three complementary subsets: a trusted information subset based on the absolute value of high-fidelity labels; a trusted subset that captures the reliability of the high-fidelity and medium-fidelity label intervals through adaptive constraints; and an ordered trusted subset representing the numerical relationships within the same fidelity level. These subsets are then integrated into a unified end-to-end model, enabling the reasonable utilization of medium- and low-fidelity information. Extensive experiments on various molecular and material property benchmarks demonstrate that our method consistently outperforms state-of-the-art multifidelity and singlefidelity baseline methods, and exhibits good robustness under real-world unpaired multifidelity conditions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5eeb3f7f-a1e2-4004-a421-565b9177b9cbBuilds on10
- Multi-Fidelity Bayesian Optimization via Deep Neural NetworksShibo Li, Wei W. Xing, Robert M. Kirby, Shandian ZheNeurIPS 2020 · 74 citations
- From Molecules to Materials: Pre-training Large Generalizable Models for Atomic Property PredictionNima Shoghi, Adeesh Kolluru, John R. Kitchin, Zachary W. Ulissi et al.ICLR 2024 · 63 citations
- MFES-HB: Efficient Hyperband with Multi-Fidelity Quality MeasurementsYang Li, Yu Shen, Jiawei Jiang, Jinyang Gao et al.AAAI 2021 · 32 citations
- RankMatch: A Novel Approach to Semi-Supervised Label Distribution Learning Leveraging Rank Correlation between LabelsZhiqiang Kou, Yucheng Xie, Hailin Wang, Junyang Chen et al.NeurIPS 2025 · 18 citations
- FedHarmony: Harmonizing Heterogeneous Label Correlations in Federated Multi-Label LearningZhiqiang Kou, Junxiang Wu, Wenke Huang, Wenwen He et al.CVPR 2026 · 3 citations
Related papers
- Rethinking Crystal Symmetry Prediction: A Decoupled PerspectiveLiheng Yu, Zhe Zhao, Xucong Wang, Di Wu et al.AAAI 2026
- Fuzzy Multimodal Learning for Trusted Cross-modal RetrievalSiyuan Duan, Yuan Sun, Dezhong Peng, Zheng Liu et al.CVPR 2025
- Physical Consistency Bridges Heterogeneous Data in Molecular Multi-Task LearningYuxuan Ren, Dihan Zheng, Chang Liu, Peiran Jin et al.NeurIPS 2024 · 3 citations
- SymSpectra: Symmetric Information Bottleneck Framework for Molecular Structure Recognition under Imbalanced SettingsXiaohan Qin, Wenjie Du, Yang WangICML 2026
- Multi-Fidelity Covariance Estimation in the Log-Euclidean GeometryAimee Maurais, Terrence Alsup, Benjamin Peherstorfer, Youssef M. MarzoukICML 2023 · 10 citations
