Defining Neural Network Architecture through Polytope Structures of Datasets
Sangmin Lee, Abbas Mammadov, Jong Chul Ye
Abstract
Current theoretical and empirical research in neural networks suggests that complex datasets require large network architectures for thorough classification, yet the precise nature of this relationship remains unclear. This paper tackles this issue by defining upper and lower bounds for neural network widths, which are informed by the polytope structure of the dataset in question. We also delve into the application of these principles to simplicial complexes and specific manifold shapes, explaining how the requirement for network width varies in accordance with the geometric complexity of the dataset. Moreover, we develop an algorithm to investigate a converse situation where the polytope structure of a dataset can be inferred from its corresponding trained neural networks. Through our algorithm, it is established that popular datasets such as MNIST, Fashion-MNIST, and CIFAR10 can be efficiently encapsulated using no more than two polytopes with a small number of faces.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5eb85796-ece5-447f-bb52-3a47c7ec8c1cCited by top-tier papers1
Ask how each one uses itBuilds on16
- Extracting Training Data from Large Language ModelsNicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski et al.USENIX Security 2021 · 2,866 citations
- The Secret Sharer: Evaluating and Testing Unintended Memorization in Neural NetworksNicholas Carlini, Chang Liu, Úlfar Erlingsson, Jernej Kos et al.USENIX Security 2019 · 1,386 citations
- Proving the Lottery Ticket Hypothesis: Pruning is All You NeedEran Malach, Gilad Yehudai, Shai Shalev-Shwartz, Ohad ShamirICML 2020 · 327 citations
- Optimal transport mapping via input convex neural networksAshok Vardhan Makkuva, Amirhossein Taghvaei, Sewoong Oh, Jason D. LeeICML 2020 · 254 citations
- Reconstructing Training Data From Trained Neural NetworksNiv Haim, Gal Vardi, Gilad Yehudai, Ohad Shamir et al.NeurIPS 2022 · 196 citations
Related papers
- Effects of Data Geometry in Early Deep LearningSaket Tiwari, George KonidarisNeurIPS 2022 · 14 citations
- Besov Function Approximation and Binary Classification on Low-Dimensional Manifolds Using Convolutional Residual NetworksHao Liu, Minshuo Chen, Tuo Zhao, Wenjing LiaoICML 2021 · 42 citations
- Characterizing the Discrete Geometry of ReLU NetworksBlake Gaines, Jinbo BiICLR 2026 · 2 citations
- Polyhedral Complex Extraction from ReLU Networks using Edge SubdivisionArturs BerzinsICML 2023 · 12 citations
- MANTRA: The Manifold Triangulations AssemblageRubén Ballester, Ernst Röell, Daniel Bin Schmid, Mathieu Alain et al.ICLR 2025
