Feature Shift Localization Network
Míriam Barrabés, Daniel Mas Montserrat, Kapal Dev, Alexander G. Ioannidis
Abstract
Feature shifts between data sources are present in many applications involving healthcare, biomedical, socioeconomic, financial, survey, and multisensor data, among others, where unharmonized heterogeneous data sources, noisy data measurements, or inconsistent processing and standardization pipelines can lead to erroneous features. Localizing shifted features is important to address the underlying cause of the shift and correct or filter the data to avoid degrading downstream analysis. While many techniques can detect distribution shifts, localizing the features originating them is still challenging, with current solutions being either inaccurate or not scalable to large and high-dimensional datasets. In this work, we introduce the Feature Shift Localization Network (FSL-Net), a neural network that can localize feature shifts in large and highdimensional datasets in a fast and accurate manner. The network, trained with a large number of datasets, learns to extract the statistical properties of the datasets and can localize feature shifts from previously unseen datasets and shifts without the need for re-training. The code and readyto-use trained model are available at https: //github.com/AI-sandbox/FSL-Net .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on7
- Clifford Group Equivariant Neural NetworksDavid Ruhe, Johannes Brandstetter, Patrick ForréNeurIPS 2023 · 85 citations
- Learning Invariances in Neural Networks from Training DataGregory W. Benton, Marc Finzi, Pavel Izmailov, Andrew Gordon WilsonNeurIPS 2020 · 78 citations
- Feature Shift Detection: Localizing Which Features Have Shifted via Conditional Distribution TestsSean Kulinski, Saurabh Bagchi, David I. InouyeNeurIPS 2020 · 39 citations
- Interactive Weak Supervision: Learning Useful Heuristics for Data LabelingBenedikt Boecking, Willie Neiswanger, Eric P. Xing, Artur DubrawskiICLR 2021 · 8 citations
- Adversarial Learning for Feature Shift Detection and CorrectionMíriam Barrabés, Daniel Mas Montserrat, Margarita Geleta, Xavier Giró-i-Nieto et al.NeurIPS 2023 · 5 citations
Related papers
- LTF: A Label Transformation Framework for Correcting Label ShiftJiaxian Guo, Mingming Gong, Tongliang Liu, Kun Zhang et al.ICML 2020 · 43 citations
- Limitations of Post-Hoc Feature Alignment for RobustnessCollin Burns, Jacob SteinhardtCVPR 2021
- Metadata NormalizationMandy Lu, Qingyu Zhao, Jiequan Zhang, Kilian M. Pohl et al.CVPR 2021
- OoD-Bench: Quantifying and Understanding Two Dimensions of Out-of-Distribution GeneralizationNanyang Ye, Kaican Li, Haoyue Bai, Runpeng Yu et al.CVPR 2022 · 74 citations
- Uncertainty Modeling for Out-of-Distribution GeneralizationXiaotong Li, Yongxing Dai, Yixiao Ge, Jun Liu et al.ICLR 2022 · 237 citations
