Lune

NeurIPS2022Top-tier venue

Learning (Very) Simple Generative Models Is Hard

Sitan Chen, Jerry Li, Yuanzhi Li

2022Year
12Citations
6Top-tier citations

Abstract

Motivated by the recent empirical successes of deep generative models, we study the computational complexity of the following unsupervised learning problem. For an unknown neural network F:Rd→Rd′F:\mathbb{R}^d\to\mathbb{R}^{d'}, let DD be the distribution over Rd′\mathbb{R}^{d'} given by pushing the standard Gaussian N(0,Idd)\mathcal{N}(0,\textrm{Id}_d) through FF. Given i.i.d. samples from DD, the goal is to output any distribution close to DD in statistical distance. We show under the statistical query (SQ) model that no polynomial-time algorithm can solve this problem even when the output coordinates of FF are one-hidden-layer ReLU networks with log⁡(d)\log(d) neurons. Previously, the best lower bounds for this problem simply followed from lower bounds for supervised learning and required at least two hidden layers and poly(d)\mathrm{poly}(d) neurons [Daniely-Vardi '21, Chen-Gollakota-Klivans-Meka '22]. The key ingredient in our proof is an ODE-based construction of a compactly supported, piecewise-linear function ff with polynomially-bounded slopes such that the pushforward of N(0,1)\mathcal{N}(0,1) under ff matches all low-degree moments of N(0,1)\mathcal{N}(0,1).

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext e13d015b-d260-4158-ad75-d133907a00a3

Cited by top-tier papers6

Ask how each one uses it

Builds on12

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines