Hadamax Encoding: Elevating Performance in Model-Free Atari
Jacob Eeuwe Kooi, Zhao Yang, Vincent François-Lavet
摘要
Neural network architectures have a large impact in machine learning. In reinforcement learning, network architectures have remained notably simple, as changes often lead to small gains in performance. This work introduces a novel encoder architecture for pixel-based model-free reinforcement learning. The Hadamax (Hadamard max-pooling) encoder achieves state-of-the-art performance by max-pooling Hadamard products between GELU-activated parallel hidden layers. Based on the recent PQN algorithm, the Hadamax encoder achieves state-of-the-art model-free performance in the Atari-57 benchmark. Specifically, without applying any algorithmic hyperparameter modifications, Hadamax-PQN achieves an 80% performance gain over vanilla PQN and significantly surpasses Rainbow-DQN. For reproducibility, the full code is available on GitHub.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Simplicial Embeddings Improve Sample Efficiency in Actor–Critic AgentsJohan Obando-Ceron, Walter Mayor, Samuel Lavoie, Scott Fujimoto 等ICLR 2026 · 被引用 12 次
- Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement LearningGhada Sokar, Pablo Samuel CastroNeurIPS 2025 · 被引用 5 次
它引用的顶会 Paper22
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 被引用 1,261 次
- Mastering Atari with Discrete World ModelsDanijar Hafner, Timothy P. Lillicrap, Mohammad Norouzi, Jimmy BaICLR 2021 · 被引用 1,170 次
- Data-Efficient Reinforcement Learning with Self-Predictive RepresentationsMax Schwarzer, Ankesh Anand, Rishab Goel, R. Devon Hjelm 等ICLR 2021 · 被引用 399 次
- Temporal Difference Learning for Model Predictive ControlNicklas Hansen, Hao Su, Xiaolong WangICML 2022 · 被引用 388 次
- Diffusion for World Modeling: Visual Details Matter in AtariEloi Alonso, Adam Jelley, Vincent Micheli, Anssi Kanervisto 等NeurIPS 2024 · 被引用 359 次
相关 Paper
- Simplifying Deep Temporal Difference LearningMatteo Gallici, Mattie Fellows, Benjamin Ellis, Bartomeu Pou 等ICLR 2025 · 被引用 1 次
- One Encoder to Rule Them All: Representation Learning for Model-Free Visual Reinforcement Learning Using Fourier Neural OperatorsParag Dutta, Mohd Ayyoob, Shalabh Bhatnagar, Ambedkar DukkipatiICCV 2025
- Reinforcement Learning with Latent FlowWenling Shang, Xiaofei Wang, Aravind Srinivas, Aravind Rajeswaran 等NeurIPS 2021 · 被引用 30 次
- Beyond The Rainbow: High Performance Deep Reinforcement Learning on a Desktop PCTyler Clark, Mark Towers, Christine Evers, Jonathon HareICML 2025
- Learning Features with Parameter-Free LayersDongyoon Han, Young Joon Yoo, Beomyoung Kim, Byeongho HeoICLR 2022 · 被引用 9 次
