Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
Hyunwoo Lee, Hayoung Choi, Hyunju Kim
Abstract
As a neural network's depth increases, it can improve generalization performance. However, training deep networks is challenging due to gradient and signal propagation issues. To address these challenges, extensive theoretical research and various methods have been introduced. Despite these advances, effective weight initialization methods for tanh neural networks remain insufficiently investigated. This paper presents a novel weight initialization method for neural networks with tanh activation function. Based on an analysis of the fixed points of the function tanh(ax), the proposed method aims to determine values of a that mitigate activation saturation. A series of experiments on various classification datasets and physics-informed neural networks demonstrates that the proposed method outperforms Xavier initialization methods (with or without normalization) in terms of robustness across different network sizes, data efficiency, and convergence speed.
Code is available at https://github.com/1HyunwooLee/Tanh-Init.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bcf70c43-dcdb-40d2-855b-03a6876f642dCited by top-tier papers1
Ask how each one uses itBuilds on3
- Minimum Width for Universal ApproximationSejun Park, Chulhee Yun, Jaeho Lee, Jinwoo ShinICLR 2021 · 148 citations
- Challenges in Training PINNs: A Loss Landscape PerspectivePratik Rathore, Weimu Lei, Zachary Frangella, Lu Lu et al.ICML 2024 · 137 citations
- MultiAdam: Parameter-wise Scale-invariant Optimizer for Multiscale Training of Physics-informed Neural NetworksJiachen Yao, Chang Su, Zhongkai Hao, Songming Liu et al.ICML 2023 · 27 citations
Related papers
- AutoInit: Analytic Signal-Preserving Weight Initialization for Neural NetworksGarrett Bingham, Risto MiikkulainenAAAI 2023 · 6 citations
- Training Deep Spiking Neural Networks without NormalizationXinyu Shi, Zhaofei YuICML 2026
- A New Initialization to Control Gradients in Sinusoidal Neural NetworksAndrea Combette, Nelly Pustelnik, Antoine VenailleICLR 2026
- Sinusoidal Initialization, Time for a New StartAlberto Fernández-Hernández, José I. Mestre, Manuel F. Dolz, José Duato et al.NeurIPS 2025 · 5 citations
- Principled Weight Initialization for HypernetworksOscar Chang, Lampros Flokas, Hod LipsonICLR 2020 · 87 citations
