Self-Attention Amortized Distributional Projection Optimization for Sliced Wasserstein Point-Cloud Reconstruction
Khai Nguyen, Dang Nguyen, Nhat Ho
Abstract
Max sliced Wasserstein (Max-SW) distance has been widely known as a solution for less discriminative projections of sliced Wasserstein (SW) distance. In applications that have various independent pairs of probability measures, amortized projection optimization is utilized to predict the ``max"projecting directions given two input measures instead of using projected gradient ascent multiple times. Despite being efficient, Max-SW and its amortized version cannot guarantee metricity property due to the sub-optimality of the projected gradient ascent and the amortization gap. Therefore, we propose to replace Max-SW with distributional sliced Wasserstein distance with von Mises-Fisher (vMF) projecting distribution (v-DSW). Since v-DSW is a metric with any non-degenerate vMF distribution, its amortized version can guarantee the metricity when performing amortization. Furthermore, current amortized models are not permutation invariant and symmetric. To address the issue, we design amortized models based on self-attention architecture. In particular, we adopt efficient self-attention architectures to make the computation linear in the number of supports. With the two improvements, we derive self-attention amortized distributional projection optimization and show its appealing performance in point-cloud reconstruction and its downstream applications.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Energy-Based Sliced Wasserstein DistanceKhai Nguyen, Nhat HoNeurIPS 2023 · 51 citations
- Quasi-Monte Carlo for 3D Sliced WassersteinKhai Nguyen, Nicola Bariletto, Nhat HoICLR 2024 · 25 citations
- Sliced Wasserstein Estimation with Control VariatesKhai Nguyen, Nhat HoICLR 2024 · 16 citations
- Linear optimal partial transport embeddingYikun Bai, Ivan Vladimir Medri, Rocio Diaz Martin, Rana Muhammad Shahroz Khan et al.ICML 2023 · 11 citations
- Diffeomorphic Mesh Deformation via Efficient Optimal Transport for Cortical Surface ReconstructionThanh-Tung Le, Khai Nguyen, Shanlin Sun, Kun Han et al.ICLR 2024 · 9 citations
Builds on16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Transformers are RNNs: Fast Autoregressive Transformers with Linear AttentionAngelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, François FleuretICML 2020 · 2,665 citations
- VideoBERT: A Joint Model for Video and Language Representation LearningChen Sun, Austin Myers, Carl Vondrick, Kevin Murphy et al.ICCV 2019 · 1,396 citations
- PointFlow: 3D Point Cloud Generation With Continuous Normalizing FlowsGuandao Yang, Xun Huang, Zekun Hao, Ming-Yu Liu et al.ICCV 2019 · 794 citations
Related papers
- Distributional Sliced-Wasserstein and Applications to Generative ModelingKhai Nguyen, Nhat Ho, Tung Pham, Hung BuiICLR 2021 · 111 citations
- Markovian Sliced Wasserstein Distances: Beyond Independent ProjectionsKhai Nguyen, Tongzheng Ren, Nhat HoNeurIPS 2023 · 13 citations
- Augmented Sliced Wasserstein DistancesXiongjie Chen, Yongxin Yang, Yunpeng LiICLR 2022 · 23 citations
- Fourier Sliced-Wasserstein Embedding for Multisets and MeasuresTal Amir, Nadav DymICLR 2025
- Towards Marginal Fairness Sliced Wasserstein BarycenterKhai Nguyen, Hai Nguyen, Nhat HoICLR 2025
