GO Hessian for Expectation-Based Objectives
Yulai Cong, Miaoyun Zhao, Jianqiao Li, Junya Chen, Lawrence Carin
摘要
An unbiased low-variance gradient estimator, termed GO gradient, was proposed recently for expectation-based objectives E_q_γ(y) [f(y)], where the random variable (RV) y may be drawn from a stochastic computation graph (SCG) with continuous (non-reparameterizable) internal nodes and continuous/discrete leaves. Based on the GO gradient, we present for E_q_γ(y) [f(y)] an unbiased low-variance Hessian estimator, named GO Hessian, which contains the deterministic Hessian as a special case. Considering practical implementation, we reveal that the GO Hessian in expectation obeys the chain rule and is therefore easy-to-use with auto-differentiation and Hessian-vector products, enabling efficient cheap exploitation of curvature information over deep SCGs. As representative examples, we present the GO Hessian for non-reparameterizable gamma and negative binomial RVs/nodes. Leveraging the GO Hessian, we develop a new second-order method for E_q_γ(y) [f(y)], with challenging experiments conducted to verify its effectiveness and efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Storchastic: A Framework for General Stochastic Automatic DifferentiationEmile van Krieken, Jakub M. Tomczak, Annette ten TeijeNeurIPS 2021 · 被引用 19 次
- Enhance Curvature Information by Structured Stochastic Quasi-Newton MethodsMinghan Yang, Dong Xu, Hongyu Chen, Zaiwen Wen 等CVPR 2021
- Automatic Differentiation of Programs with Discrete RandomnessGaurav Arya, Moritz Schauer, Frank Schäfer, Christopher RackauckasNeurIPS 2022 · 被引用 56 次
- Generalized Doubly Reparameterized Gradient EstimatorsMatthias Bauer, Andriy MnihICML 2021 · 被引用 15 次
- Small Steps and Giant Leaps: Minimal Newton Solvers for Deep LearningJoão F. Henriques, Sébastien Ehrhardt, Samuel Albanie, Andrea VedaldiICCV 2019 · 被引用 23 次
