Lune

NeurIPS2020顶会

A Universal Approximation Theorem of Deep Neural Networks for Expressing Probability Distributions

Yulong Lu, Jianfeng Lu

2020年份
146被引次数
21顶会引用

摘要

This paper studies the universal approximation property of deep neural networks for representing probability distributions. Given a target distribution ππ and a source distribution pzp_z both defined on Rd\mathbb{R}^d, we prove under some assumptions that there exists a deep neural network g:Rd→Rg:\mathbb{R}^d\rightarrow \mathbb{R} with ReLU activation such that the push-forward measure (∇g)#pz(\nabla g)_\# p_z of pzp_z under the map ∇g\nabla g is arbitrarily close to the target measure ππ. The closeness are measured by three classes of integral probability metrics between probability distributions: 11-Wasserstein distance, maximum mean distance (MMD) and kernelized Stein discrepancy (KSD). We prove upper bounds for the size (width and depth) of the deep neural network in terms of the dimension dd and the approximation error ε\varepsilon with respect to the three discrepancies. In particular, the size of neural network can grow exponentially in dd when 11-Wasserstein distance is used as the discrepancy, whereas for both MMD and KSD the size of neural network only depends on dd at most polynomially. Our proof relies on convergence estimates of empirical measures under aforementioned discrepancies and semi-discrete optimal transport.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper21

问问它们各自怎么用它

它引用的顶会 Paper1

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖