Hadamax Encoding: Elevating Performance in Model-Free Atari
Jacob Eeuwe Kooi, Zhao Yang, Vincent François-Lavet
Abstract
Neural network architectures have a large impact in machine learning. In reinforcement learning, network architectures have remained notably simple, as changes often lead to small gains in performance. This work introduces a novel encoder architecture for pixel-based model-free reinforcement learning. The Hadamax (Hadamard max-pooling) encoder achieves state-of-the-art performance by max-pooling Hadamard products between GELU-activated parallel hidden layers. Based on the recent PQN algorithm, the Hadamax encoder achieves state-of-the-art model-free performance in the Atari-57 benchmark. Specifically, without applying any algorithmic hyperparameter modifications, Hadamax-PQN achieves an 80% performance gain over vanilla PQN and significantly surpasses Rainbow-DQN. For reproducibility, the full code is available on GitHub.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Simplicial Embeddings Improve Sample Efficiency in Actor–Critic AgentsJohan Obando-Ceron, Walter Mayor, Samuel Lavoie, Scott Fujimoto et al.ICLR 2026 · 12 citations
- Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement LearningGhada Sokar, Pablo Samuel CastroNeurIPS 2025 · 5 citations
Builds on22
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 1,261 citations
- Mastering Atari with Discrete World ModelsDanijar Hafner, Timothy P. Lillicrap, Mohammad Norouzi, Jimmy BaICLR 2021 · 1,170 citations
- Data-Efficient Reinforcement Learning with Self-Predictive RepresentationsMax Schwarzer, Ankesh Anand, Rishab Goel, R. Devon Hjelm et al.ICLR 2021 · 399 citations
- Temporal Difference Learning for Model Predictive ControlNicklas Hansen, Hao Su, Xiaolong WangICML 2022 · 388 citations
- Diffusion for World Modeling: Visual Details Matter in AtariEloi Alonso, Adam Jelley, Vincent Micheli, Anssi Kanervisto et al.NeurIPS 2024 · 359 citations
Related papers
- Simplifying Deep Temporal Difference LearningMatteo Gallici, Mattie Fellows, Benjamin Ellis, Bartomeu Pou et al.ICLR 2025 · 1 citation
- One Encoder to Rule Them All: Representation Learning for Model-Free Visual Reinforcement Learning Using Fourier Neural OperatorsParag Dutta, Mohd Ayyoob, Shalabh Bhatnagar, Ambedkar DukkipatiICCV 2025
- Reinforcement Learning with Latent FlowWenling Shang, Xiaofei Wang, Aravind Srinivas, Aravind Rajeswaran et al.NeurIPS 2021 · 30 citations
- Beyond The Rainbow: High Performance Deep Reinforcement Learning on a Desktop PCTyler Clark, Mark Towers, Christine Evers, Jonathon HareICML 2025
- Learning Features with Parameter-Free LayersDongyoon Han, Young Joon Yoo, Beomyoung Kim, Byeongho HeoICLR 2022 · 9 citations
