Using Random Effects to Account for High-Cardinality Categorical Features and Repeated Measures in Deep Neural Networks
Giora Simchoni, Saharon Rosset
摘要
High-cardinality categorical features are a major challenge for machine learning methods in general and for deep learning in particular. Existing solutions such as one-hot encoding and entity embeddings can be hard to scale when the cardinality is very high, require much space, are hard to interpret or may overfit the data. A special scenario of interest is that of repeated measures, where the categorical feature is the identity of the individual or object, and each object is measured several times, possibly under different conditions (values of the other features). We propose accounting for high-cardinality categorical features as random effects variables in a regression setting, and consequently adopt the corresponding negative log likelihood loss from the linear mixed models (LMM) statistical literature and integrate it in a deep learning framework. We test our model which we call LMMNN on simulated as well as real datasets with a single categorical feature with high cardinality, using various baseline neural networks architectures such as convolutional networks and LSTM, and various applications in e-commerce, healthcare and computer vision. Our results show that treating high-cardinality categorical features as random effects leads to a significant improvement in prediction performance compared to state of the art alternatives. Potential extensions such as accounting for multiple categorical features and classification settings are discussed. Our code and simulations are available at https://github.com/gsimchoni/lmmnn
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Label Correction of Crowdsourced Noisy Annotations with an Instance-Dependent Noise Transition ModelHui Guo, Boyu Wang, Grace YiNeurIPS 2023 · 被引用 21 次
- H-Likelihood Approach to Deep Neural Networks with Temporal-Spatial Random Effects for High-Cardinality Categorical FeaturesHangbin Lee, Youngjo LeeICML 2023 · 被引用 4 次
- Training Normalizing Flows from Dependent DataMatthias Kirchler, Christoph Lippert, Marius KloftICML 2023 · 被引用 2 次
- Hamiltonian Monte Carlo Inference of Marginalized Linear Mixed-Effects ModelsJinlin Lai, Justin Domke, Daniel R. SheldonNeurIPS 2024 · 被引用 2 次
它引用的顶会 Paper1
相关 Paper
- LMLFM: Longitudinal Multi-Level Factorization MachineJunjie Liang, Dongkuan Xu, Yiwei Sun, Vasant G. HonavarAAAI 2020 · 被引用 12 次
- How do Categorical Duplicates Affect ML? A New Benchmark and Empirical AnalysesVraj Shah, Thomas J. Parashos, Arun KumarVLDB 2024 · 被引用 8 次
- Learning to Embed Categorical Features without Embedding Tables for RecommendationWang-Cheng Kang, Derek Zhiyuan Cheng, Tiansheng Yao, Xinyang Yi 等KDD 2021 · 被引用 46 次
- Unified Embedding: Battle-Tested Feature Representations for Web-Scale ML SystemsBenjamin Coleman, Wang-Cheng Kang, Matthew Fahrbach, Ruoxi Wang 等NeurIPS 2023 · 被引用 30 次
- Field-wise Learning for Multi-field Categorical DataZhibin Li, Jian Zhang, Yongshun Gong, Yazhou Yao 等NeurIPS 2020 · 被引用 10 次
