Layered-Parameter Perturbation for Zeroth-Order Optimization of Optical Neural Networks
Hiroshi Sawada, Kazuo Aoyama, Masaya Notomi
Abstract
Optical neural networks (ONNs) have attracted great attention due to their low power consumption and high-speed processing. When training an ONN implemented on a chip with possible fabrication variations, the well-known backpropagation algorithm cannot be executed accurately because the perfect information inside the chip cannot be observed. Instead, we employ a black-box optimization method such as zeroth-order (ZO) optimization. In this paper, we first discuss how ONN parameters should be perturbed to search for better values in a black-box manner. Conventionally, parameter perturbations are sampled from a normal distribution with an identity covariance matrix. This is plausible if the parameters are not interrelated in a module, like a linear module of an ordinary neural network. However, this is not the best way for ONN modules with layered parameters, which are interrelated by optical paths. We then propose to perturb the parameters by a normal distribution with a special covariance matrix computed by our novel method. The covariance matrix is designed so that the perturbations appearing at the module output caused by the parameter perturbations become as isotropic as possible to uniformly search for better values. Experimental results show that the proposed method using the special covariance matrix significantly outperformed conventional methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on3
- Efficient On-Chip Learning for Optical Neural Networks Through Power-Aware Sparse Zeroth-Order OptimizationJiaqi Gu, Chenghao Feng, Zheng Zhao, Zhoufeng Ying et al.AAAI 2021 · 41 citations
- FLOPS: EFficient On-Chip Learning for OPtical Neural Networks Through Stochastic Zeroth-Order OptimizationJiaqi Gu, Zheng Zhao, Chenghao Feng, Wuxi Li et al.DAC 2020 · 20 citations
- Zeroth-Order Optimization of Optical Neural Networks with Linear Combination Natural Gradient and Calibrated ModelHiroshi Sawada, Kazuo Aoyama, Kohei IkedaDAC 2024 · 1 citation
Related papers
- Learning to Learn by Zeroth-Order OracleYangjun Ruan, Yuanhao Xiong, Sashank J. Reddi, Sanjiv Kumar et al.ICLR 2020 · 21 citations
- Zeroth-Order Forward-Only SNN Training Inspiring Neuromorphic On-Chip LearningMingyue Qin, Shuyu Yin, Qinghai Guo, Peilin Liu et al.ICML 2026
- Binary Optical Machine Learning: Million-Scale Physical Neural Networks with Nano NeuronsXueyuan Yang, Zhenlin An, Qingrui Pan, Lei Yang et al.MobiCom 2024 · 3 citations
- Power-aware pruning for ultrafast, energy-efficient, and accurate optical neural network designNaoki Hattori, Yutaka Masuda, Tohru Ishihara, Akihiko Shinya et al.DAC 2022 · 1 citation
- Towards Query-Efficient Black-Box Adversary with Zeroth-Order Natural Gradient DescentPu Zhao, Pin-Yu Chen, Siyue Wang, Xue LinAAAI 2020 · 42 citations
