FiLM-Ensemble: Probabilistic Deep Learning via Feature-wise Linear Modulation
Mehmet Ozgur Turkoglu, Alexander Becker, Hüseyin Anil Gündüz, Mina Rezaei, Bernd Bischl, Rodrigo Caye Daudt, Stefano D'Aronco, Jan D. Wegner, Konrad Schindler
摘要
The ability to estimate epistemic uncertainty is often crucial when deploying machine learning in the real world, but modern methods often produce overconfident, uncalibrated uncertainty predictions. A common approach to quantify epistemic uncertainty, usable across a wide class of prediction models, is to train a model ensemble. In a naive implementation, the ensemble approach has high computational cost and high memory demand. This challenges in particular modern deep learning, where even a single deep network is already demanding in terms of compute and memory, and has given rise to a number of attempts to emulate the model ensemble without actually instantiating separate ensemble members. We introduce FiLM-Ensemble, a deep, implicit ensemble method based on the concept of Feature-wise Linear Modulation (FiLM). That technique was originally developed for multi-task learning, with the aim of decoupling different tasks. We show that the idea can be extended to uncertainty quantification: by modulating the network activations of a single deep network with FiLM, one obtains a model ensemble with high diversity, and consequently well-calibrated estimates of epistemic uncertainty, with low computational overhead in comparison. Empirically, FiLM-Ensemble outperforms other implicit ensemble methods, and it and comes very close to the upper bound of an explicit ensemble of networks (sometimes even beating it), at a fraction of the memory cost.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Improving day-ahead Solar Irradiance Time Series Forecasting by Leveraging Spatio-Temporal ContextOussama Boussif, Ghait Boukachab, Dan Assouline, Stefano Massaroli 等NeurIPS 2023 · 被引用 46 次
- Variational Multi-scale Representation for Estimating Uncertainty in 3D Gaussian SplattingRuiqi Li, Yiu-ming CheungNeurIPS 2024 · 被引用 26 次
- Understanding Multi-Granularity for Open-Vocabulary Part SegmentationJiho Choi, Seonho Lee, Seungho Lee, Minhyun Lee 等NeurIPS 2024 · 被引用 7 次
- Regularized Multi-Decoder Ensemble for an Error-Aware Scene Representation NetworkTianyu Xiong, Skylar W. Wurster, Hanqi Guo, Tom Peterka 等IEEE VIS 2024 · 被引用 6 次
- Split-Ensemble: Efficient OOD-aware Ensemble via Task and Model SplittingAnthony Chen, Huanrui Yang, Yulu Gan, Denis A. Gudovskiy 等ICML 2024 · 被引用 5 次
它引用的顶会 Paper10
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 被引用 845 次
- Simple and Principled Uncertainty Estimation with Deterministic Deep Learning via Distance AwarenessJeremiah Z. Liu, Zi Lin, Shreyas Padhy, Dustin Tran 等NeurIPS 2020 · 被引用 604 次
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 被引用 569 次
- What Are Bayesian Neural Network Posteriors Really Like?Pavel Izmailov, Sharad Vikram, Matthew D. Hoffman, Andrew Gordon WilsonICML 2021 · 被引用 458 次
相关 Paper
- Quantifying the Uncertainty of Foundation Models with Singular Value EnsemblesMehmet Ozgur Turkoglu, Dominik J. Mühlematter, Alexander Becker, Konrad Schindler 等ICML 2026
- Epistemic Neural NetworksIan Osband, Zheng Wen, Seyed Mohammad Asghari, Vikranth Dwaracherla 等NeurIPS 2023 · 被引用 142 次
- Is the Last Layer Sufficient for Uncertainty Quantification?Joseph Wilson, Chris van der Heide, Liam Hodgkinson, Fred RoostaICML 2026
- Evidential Turing ProcessesMelih Kandemir, Abdullah Akgül, Manuel Haußmann, Gozde UnalICLR 2022 · 被引用 10 次
- Hyperdimensional Uncertainty Quantification for Multimodal Uncertainty Fusion in Autonomous Vehicles PerceptionLuke Chen, Junyao Wang, Trier Mortlock, Pramod P. Khargonekar 等CVPR 2025
