Layered-Parameter Perturbation for Zeroth-Order Optimization of Optical Neural Networks
Hiroshi Sawada, Kazuo Aoyama, Masaya Notomi
摘要
Optical neural networks (ONNs) have attracted great attention due to their low power consumption and high-speed processing. When training an ONN implemented on a chip with possible fabrication variations, the well-known backpropagation algorithm cannot be executed accurately because the perfect information inside the chip cannot be observed. Instead, we employ a black-box optimization method such as zeroth-order (ZO) optimization. In this paper, we first discuss how ONN parameters should be perturbed to search for better values in a black-box manner. Conventionally, parameter perturbations are sampled from a normal distribution with an identity covariance matrix. This is plausible if the parameters are not interrelated in a module, like a linear module of an ordinary neural network. However, this is not the best way for ONN modules with layered parameters, which are interrelated by optical paths. We then propose to perturb the parameters by a normal distribution with a special covariance matrix computed by our novel method. The covariance matrix is designed so that the perturbations appearing at the module output caused by the parameter perturbations become as isotropic as possible to uniformly search for better values. Experimental results show that the proposed method using the special covariance matrix significantly outperformed conventional methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- Efficient On-Chip Learning for Optical Neural Networks Through Power-Aware Sparse Zeroth-Order OptimizationJiaqi Gu, Chenghao Feng, Zheng Zhao, Zhoufeng Ying 等AAAI 2021 · 被引用 41 次
- FLOPS: EFficient On-Chip Learning for OPtical Neural Networks Through Stochastic Zeroth-Order OptimizationJiaqi Gu, Zheng Zhao, Chenghao Feng, Wuxi Li 等DAC 2020 · 被引用 20 次
- Zeroth-Order Optimization of Optical Neural Networks with Linear Combination Natural Gradient and Calibrated ModelHiroshi Sawada, Kazuo Aoyama, Kohei IkedaDAC 2024 · 被引用 1 次
相关 Paper
- Learning to Learn by Zeroth-Order OracleYangjun Ruan, Yuanhao Xiong, Sashank J. Reddi, Sanjiv Kumar 等ICLR 2020 · 被引用 21 次
- Zeroth-Order Forward-Only SNN Training Inspiring Neuromorphic On-Chip LearningMingyue Qin, Shuyu Yin, Qinghai Guo, Peilin Liu 等ICML 2026
- Binary Optical Machine Learning: Million-Scale Physical Neural Networks with Nano NeuronsXueyuan Yang, Zhenlin An, Qingrui Pan, Lei Yang 等MobiCom 2024 · 被引用 3 次
- Power-aware pruning for ultrafast, energy-efficient, and accurate optical neural network designNaoki Hattori, Yutaka Masuda, Tohru Ishihara, Akihiko Shinya 等DAC 2022 · 被引用 1 次
- Towards Query-Efficient Black-Box Adversary with Zeroth-Order Natural Gradient DescentPu Zhao, Pin-Yu Chen, Siyue Wang, Xue LinAAAI 2020 · 被引用 42 次
