LP-SparseMAP: Differentiable Relaxed Optimization for Sparse Structured Prediction
Vlad Niculae, André F. T. Martins
Abstract
Structured predictors require solving a combinatorial optimization problem over a large number of structures, such as dependency trees or alignments. When embedded as structured hidden layers in a neural net, argmin differentiation and efficient gradient computation are further required. Recently, SparseMAP has been proposed as a differentiable, sparse alternative to maximum a posteriori (MAP) and marginal inference. SparseMAP returns an interpretable combination of a small number of structures; its sparsity being the key to efficient optimization. However, SparseMAP requires access to an exact MAP oracle in the structured model, excluding, e.g., loopy graphical models or logic constraints, which generally require approximate inference. In this paper, we introduce LP-SparseMAP, an extension of SparseMAP addressing this limitation via a local polytope relaxation. LP-SparseMAP uses the flexible and powerful language of factor graphs to define expressive hidden structures, supporting coarse decompositions, hard logic constraints, and higher-order correlations. We derive the forward and backward algorithms needed for using LP-SparseMAP as a structured hidden or output layer. Experiments in three structured tasks show benefits versus SparseMAP and Structured SVM.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d84d046d-f69c-4682-afd1-eab69e4bad72Cited by top-tier papers8
- Efficient and Modular Implicit DifferentiationMathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig et al.NeurIPS 2022 · 386 citations
- Implicit MLE: Backpropagating Through Discrete Exponential Family DistributionsMathias Niepert, Pasquale Minervini, Luca FranceschiNeurIPS 2021 · 121 citations
- Adaptive Perturbation-Based Gradient Estimation for Discrete Latent Variable ModelsPasquale Minervini, Luca Franceschi, Mathias NiepertAAAI 2023 · 16 citations
- Learning Discrete Structured Variational Auto-Encoder using Natural Evolution StrategiesAlon Berliner, Guy Rotman, Yossi Adi, Roi Reichart et al.ICLR 2022 · 5 citations
- Improved Latent Tree Induction with Distant Supervision via Span ConstraintsZhiyang Xu, Andrew Drozdov, Jay-Yoon Lee, Tim O'Gorman et al.EMNLP 2021 · 4 citations
Related papers
- Semantic Probabilistic Layers for Neuro-Symbolic LearningKareem Ahmed, Stefano Teso, Kai-Wei Chang, Guy Van den Broeck et al.NeurIPS 2022 · 133 citations
- LogicMP: A Neuro-symbolic Approach for Encoding First-order Logic ConstraintsWeidi Xu, Jingwei Wang, Lele Xie, Jianshan He et al.ICLR 2024 · 6 citations
- Efficient Marginalization of Discrete and Structured Latent Variables via SparsityGonçalo M. Correia, Vlad Niculae, Wilker Aziz, André F. T. MartinsNeurIPS 2020 · 25 citations
- Modeling Structure with Undirected Neural NetworksTsvetomila Mihaylova, Vlad Niculae, André F. T. MartinsICML 2022 · 1 citation
- When Logic Meets Perception: Operator-Agnostic Differentiable Reasoning for Reliable Neural PredictionZihan Shao, Chang Lu, Renate A. Schmidt, Yizheng ZhaoKDD 2026
