Heterogeneous Multi-Agent Reinforcement Learning with Attention for Cooperative and Scalable Feature Transformation
Tao Zhe, Huazhen Fang, Kunpeng Liu, Qian Lou, Tamzidul Hoque, Dongjie Wang
摘要
Feature transformation enhances downstream task performance by generating informative features through mathematical feature crossing. Despite the advancements in deep learning, feature transformation remains essential, particularly for structured data, where deep models often struggle to capture complex feature interactions effectively. Prior literature on automated feature transformation has achieved notable success but often relies on heuristics or exhaustive searches, leading to inefficient and time-consuming processes. Recent works employ reinforcement learning (RL) to enhance traditional approaches through a more effective trial-and-error way. However, two key limitations remain: 1) Dynamic feature expansion during the transformation process, which introduces instability and increases the time complexity of the learning procedure for RL agents; 2) Insufficient cooperation and communication between agents, which results in suboptimal feature crossing operations and degraded model performance. To address them, we propose a novel heterogeneous multi-agent RL framework to enable cooperative and scalable feature transformation. The framework comprises three heterogeneous agents, grouped into two types, each designed to select essential features and operations for feature crossing. To enhance communication among these agents, we implement a shared critic mechanism that facilitates information exchange during the feature transformation process. This collaboration enables the agents to learn more intelligent and effective transformation policies. To handle the dynamically expanding feature space, we tailor multi-head attention-based feature agents to select suitable features for feature crossing. This design facilitates scalable decision-making and effective candidate selection based on comprehensive global feature space information. Additionally, we introduce a state encoding technique during the optimization process to stabilize and enhance the learning dynamics of the RL agents, resulting in more robust and reliable transformation policies. Finally, we conduct extensive experiments to validate the effectiveness, efficiency, robustness, and interpretability of our model. Our code and dataset are publicly available on GitHub.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Multi-Agent Reinforcement Learning is a Sequence Modeling ProblemMuning Wen, Jakub Grudzien Kuba, Runji Lin, Weinan Zhang 等NeurIPS 2022 · 被引用 408 次
- Trust Region Policy Optimisation in Multi-Agent Reinforcement LearningJakub Grudzien Kuba, Ruiqing Chen, Muning Wen, Ying Wen 等ICLR 2022 · 被引用 367 次
- Reinforcement-Enhanced Autoregressive Feature Transformation: Gradient-steered Search in Continuous Space for Postfix ExpressionsDongjie Wang, Meng Xiao, Min Wu, Pengfei Wang 等NeurIPS 2023 · 被引用 34 次
- Group-wise Reinforcement Feature Generation for Optimal and Explainable Representation Space ReconstructionDongjie Wang, Yanjie Fu, Kunpeng Liu, Xiaolin Li 等KDD 2022 · 被引用 26 次
- Sequential Asynchronous Action Coordination in Multi-Agent Systems: A Stackelberg Decision Transformer ApproachBin Zhang, Hangyu Mao, Lijuan Li, Zhiwei Xu 等ICML 2024 · 被引用 13 次
相关 Paper
- Fastft: Accelerating Reinforced Feature Transformation via Advanced Exploration StrategiesTianqi He, Xiaohan Huang, Yi Du, Qingqing Long 等ICDE 2025 · 被引用 4 次
- Evolutionary Large Language Model for Automated Feature TransformationNanxu Gong, Chandan K. Reddy, Wangyang Ying, Haifeng Chen 等AAAI 2025 · 被引用 39 次
- MORE-FE: Multi-Operator and Reinforcement Learning-Enhanced Evolution for LLM Feature EngineeringChang-Yu Chao, Bryan Andersen, Xiao Xi Tan, Yi-Tse Lu 等KDD 2026
- Heterogeneous Skill Learning for Multi-agent TasksYuntao Liu, Yuan Li, Xinhai Xu, Yong Dou 等NeurIPS 2022 · 被引用 33 次
- UPDeT: Universal Multi-agent RL via Policy Decoupling with TransformersSiyi Hu, Fengda Zhu, Xiaojun Chang, Xiaodan LiangICLR 2021 · 被引用 49 次
