Catch: Collaborative Feature Set Search for Automated Feature Engineering
Guoshan Lu, Haobo Wang, Saisai Yang, Jing Yuan, Guozheng Yang, Cheng Zang, Gang Chen, Junbo Zhao
摘要
Feature engineering often plays a crucial role in building mining systems for tabular data, which traditionally requires experienced human experts to perform. Thanks to the rapid advances in reinforcement learning, it has offered an automated alternative, i.e. automated feature engineering (AutoFE). In this work, through scrutiny of the prior AutoFE methods, we characterize several research challenges that remained in this regime, concerning system-wide efficiency, efficacy, and practicality toward production. We then propose Catch, a full-fledged new AutoFE framework that comprehensively addresses the aforementioned challenges. The core to Catch composes a hierarchical-policy reinforcement learning scheme that manifests a collaborative feature engineering exploration and exploitation grounded on the granularity of the whole feature set. At a higher level of the hierarchy, a decision-making module controls the post-processing of the attained feature engineering transformation. We extensively experiment with Catch on 26 academic standardized tabular datasets and 9 industrialized real-world datasets. Measured by numerous metrics and analyses, Catch establishes a new state-of-the-art, from perspectives performance, latency as well as its practicality towards production. Source code1 can be found at https://github.com/1171000709/Catch.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Learning a Data-Driven Policy Network for Pre-Training Automated Feature EngineeringLiyao Li, Haobo Wang, Liangyu Zha, Qingyi Huang 等ICLR 2023
- Toward Efficient Automated Feature EngineeringKafeng Wang, Pengyang Wang, Chengzhong XuICDE 2023 · 被引用 6 次
- CoFE: Collaborative Feature Engineering via Semantically-Guided Exploration and Diagnostic-Driven RefinementWeihao Jiang, Ziang Nan, Zhihui Shi, Ya Cong 等KDD 2026
- OpenFE: Automated Feature Generation with Expert-level PerformanceTianping Zhang, Zheyu Aqa Zhang, Zhiyuan Fan, Haoyan Luo 等ICML 2023 · 被引用 60 次
- Group-wise Reinforcement Feature Generation for Optimal and Explainable Representation Space ReconstructionDongjie Wang, Yanjie Fu, Kunpeng Liu, Xiaolin Li 等KDD 2022 · 被引用 26 次
