Analytical-Chemistry-Informed Transformer for Infrared Spectra Modeling
Shiluo Huang, Yining Jin, Wei Jin, Ying Mu
Abstract
Infrared (IR) spectroscopy is a fundamental technique in analytical chemistry. Recently, deep learning (DL) has drawn great interest as the modeling method of infrared spectral data. However, unlike vision or language tasks, IR spectral data modeling is faced with the problem of calibration transfer and has distinctive characteristics. Introducing the prior knowledge of IR spectroscopy could guide the DL methods to learn representations aligned with the domain-invariant characteristics of spectra, and thus improve the performance. Despite such potential, there is a notable absence of DL methods that incorporate such inductive bias. To this end, we propose Analytical-Chemistry-Informed Transformer (ACT) with two modules informed by the field knowledge in analytical chemistry. First, ACT includes learnable spectral processing inspired by chemometrics, which comprises spectral pre-processing, tokenization, and post-processing. Second, a straightforward yet effective representation learning mechanism, namely spectral-attention, is incorporated into ACT. Spectral-attention utilizes the intra-spectral and inter-spectral correlations to extract spectral representations. Empirical results show that ACT has achieved competitive results in 9 analytical tasks covering applications across pharmacy, chemistry, and agriculture. Compared with existing networks, ACT reduces the root mean square error of prediction (RMSEP) by more than 20% in calibration transfer tasks. These results indicate that DL methods in IR spectroscopy could benefit from the integration of prior knowledge in analytical chemistry.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on3
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 5,824 citations
- A Time Series is Worth 64 Words: Long-term Forecasting with TransformersYuqi Nie, Nam H. Nguyen, Phanwadee Sinthong, Jayant KalagnanamICLR 2023 · 536 citations
- Adaptive Normalization for Non-stationary Time Series Forecasting: A Temporal Slice PerspectiveZhiding Liu, Mingyue Cheng, Zhi Li, Zhenya Huang et al.NeurIPS 2023 · 162 citations
Related papers
- SpectraLLM: Uncovering the Ability of LLMs for Molecule Structure Elucidation from Multi-SpectraYunyue Su, Jiahui Chen, Zao Jiang, Zhenyi Zhong et al.ICLR 2026 · 1 citation
- GTMGC: Using Graph Transformer to Predict Molecule's Ground-State ConformationGuikun Xu, Yongquan Jiang, PengChuan Lei, Yan Yang et al.ICLR 2024 · 7 citations
- CARL: Camera-Agnostic Representation Learning for Spectral Image AnalysisAlexander Baumann, Leonardo Ayala, Silvia Seidlitz, Jan Sellner et al.ICLR 2026 · 2 citations
- MolTailor: Tailoring Chemical Molecular Representation to Specific Tasks via Text PromptsHaoqiang Guo, Sendong Zhao, Haochun Wang, Yanrui Du et al.AAAI 2024 · 17 citations
- Rethinking Graph Transformers with Spectral AttentionDevin Kreuzer, Dominique Beaini, William L. Hamilton, Vincent Létourneau et al.NeurIPS 2021 · 854 citations
