Automatic Differentiation of Programs with Discrete Randomness
Gaurav Arya, Moritz Schauer, Frank Schäfer, Christopher Rackauckas
摘要
Automatic differentiation (AD), a technique for constructing new programs which compute the derivative of an original program, has become ubiquitous throughout scientific computing and deep learning due to the improved performance afforded by gradient-based optimization. However, AD systems have been restricted to the subset of programs that have a continuous dependence on parameters. Programs that have discrete stochastic behaviors governed by distribution parameters, such as flipping a coin with probability of being heads, pose a challenge to these systems because the connection between the result (heads vs tails) and the parameters () is fundamentally discrete. In this paper we develop a new reparameterization-based methodology that allows for generating programs whose expectation is the derivative of the expectation of the original program. We showcase how this method gives an unbiased and low-variance estimator which is as automated as traditional AD mechanisms. We demonstrate unbiased forward-mode AD of discrete-time Markov chains, agent-based models such as Conway's Game of Life, and unbiased reverse-mode AD of a particle filter. Our code package is available at https://github.com/gaurav-arya/StochasticAD.jl.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- ADEV: Sound Automatic Differentiation of Expected Values of Probabilistic ProgramsAlexander K. Lew, Mathieu Huot, Sam Staton, Vikash K. MansinghkaPOPL 2023 · 被引用 16 次
- Probabilistic Programming with Programmable Variational InferenceMcCoy R. Becker, Alexander K. Lew, Xiaoyan Wang, Matin Ghavami 等PLDI 2024 · 被引用 8 次
- Auto-Differentiation of Relational Computations for Very Large Scale Machine LearningYuxin Tang, Zhimin Ding, Dimitrije Jankov, Binhang Yuan 等ICML 2023 · 被引用 7 次
- Learning Individual Behavior in Agent-Based Models with Graph Diffusion NetworksFrancesco Cozzi, Marco Pangallo, Alan Perotti, André Panisson 等NeurIPS 2025 · 被引用 5 次
- Generalizing Stochastic Smoothing for Differentiation and Gradient EstimationFelix Petersen, Christian Borgelt, Aashwin Mishra, Stefano ErmonICML 2026 · 被引用 4 次
它引用的顶会 Paper4
- Differentiable Particle Filtering via Entropy-Regularized Optimal TransportAdrien Corenflos, James Thornton, George Deligiannidis, Arnaud DoucetICML 2021 · 被引用 91 次
- Storchastic: A Framework for General Stochastic Automatic DifferentiationEmile van Krieken, Jakub M. Tomczak, Annette ten TeijeNeurIPS 2021 · 被引用 19 次
- ADEV: Sound Automatic Differentiation of Expected Values of Probabilistic ProgramsAlexander K. Lew, Mathieu Huot, Sam Staton, Vikash K. MansinghkaPOPL 2023 · 被引用 16 次
- Direct Policy Gradients: Direct Optimization of Policies in Discrete Action SpacesGuy Lorberbom, Chris J. Maddison, Nicolas Heess, Tamir Hazan 等NeurIPS 2020 · 被引用 8 次
相关 Paper
- Randomized Automatic DifferentiationDeniz Oktay, Nick McGreivy, Joshua Aduol, Alex Beatson 等ICLR 2021 · 被引用 31 次
- Fiber Monte CarloNick Richardson, Deniz Oktay, Yaniv Ovadia, James C. Bowden 等ICLR 2024
- GO Hessian for Expectation-Based ObjectivesYulai Cong, Miaoyun Zhao, Jianqiao Li, Junya Chen 等AAAI 2021
- Diffusion Differentiable ResamplingJennifer R. Andersson, Zheng ZhaoICML 2026 · 被引用 2 次
- Slice Sampling Reparameterization GradientsDavid M. Zoltowski, Diana Cai, Ryan P. AdamsNeurIPS 2021 · 被引用 8 次
