Modeling Label Space Interactions in Multi-label Classification using Box Embeddings
Dhruvesh Patel, Pavitra Dangati, Jay-Yoon Lee, Michael Boratko, Andrew McCallum
摘要
Multi-label classification is a challenging structured prediction task in which a set of output class labels are predicted for each input. Real-world datasets often have natural or latent taxonomic relationships between labels, making it desirable for models to employ label representations capable of capturing such taxonomies. Most existing multi-label classification methods do not do so, resulting in label predictions that are inconsistent with the taxonomic constraints, thus failing to accurately represent the fundamentals of problem setting. In this work we introduce the multi-label box model (MBM), a multi-label classification method that combines the encoding power of neural networks with the inductive bias and probabilistic semantics of box embeddings (Vilnis, et al 2018). Box embeddings can be understood as trainable Venn-diagrams based on hyper-rectangles. Representing labels by boxes rather than vectors, MBM is able to capture taxonomic relations among labels. Furthermore, since box embeddings allow these relations to be learned by stochastic gradient descent from data, and to be read as calibrated conditional probabilities, our model is endowed with a high degree of interpretability. This interpretability also facilitates the injection of partial information about label-label relationships into model training, to further improve its consistency. We provide theoretical grounding for our method and show experimentally the model's ability to learn the true latent taxonomic structure from data. Through extensive empirical evaluations on both small and large-scale multi-label classification datasets, we show that BBM can significantly improve taxonomic consistency while preserving or surpassing state-of-the-art predictive performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Structured Prediction with Stronger Consistency GuaranteesAnqi Mao, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 被引用 37 次
- Hyperbolic Embedding Inference for Structured Multi-Label PredictionBo Xiong, Michael Cochez, Mojtaba Nayyeri, Steffen StaabNeurIPS 2022 · 被引用 23 次
- Don't Pour Cereal into Coffee: Differentiable Temporal Logic for Temporal Action SegmentationZiwei Xu, Yogesh S. Rawat, Yongkang Wong, Mohan S. Kankanhalli 等NeurIPS 2022 · 被引用 18 次
- GammaE: Gamma Embeddings for Logical Queries on Knowledge GraphsDong Yang, Peijun Qing, Yang Li, Haonan Lu 等EMNLP 2022 · 被引用 14 次
- Self-Paced Unified Representation Learning for Hierarchical Multi-Label ClassificationZixuan Yuan, Hao Liu, Haoyi Zhou, Denghui Zhang 等AAAI 2024 · 被引用 4 次
它引用的顶会 Paper4
- Hyperbolic Neural Networks++Ryohei Shimizu, Yusuke Mukuta, Tatsuya HaradaICLR 2021 · 被引用 791 次
- Coherent Hierarchical Multi-Label Classification NetworksEleonora Giunchiglia, Thomas LukasiewiczNeurIPS 2020 · 被引用 142 次
- Improving Local Identifiability in Probabilistic Box EmbeddingsShib Sankar Dasgupta, Michael Boratko, Dongxu Zhang, Luke Vilnis 等NeurIPS 2020 · 被引用 75 次
- Modeling Fine-Grained Entity Types with Box EmbeddingsYasumasa Onoe, Michael Boratko, Andrew McCallum, Greg DurrettACL 2021
相关 Paper
- A Single Vector Is Not Enough: Taxonomy Expansion via Box EmbeddingsSong Jiang, Qiyue Yao, Qifan Wang, Yizhou SunWWW 2023 · 被引用 20 次
- Hyperbolic Interaction Model for Hierarchical Multi-Label ClassificationBoli Chen, Xin Huang, Lin Xiao, Zixin Cai 等AAAI 2020 · 被引用 78 次
- TaxoBell: Gaussian Box Embeddings for Self-Supervised Taxonomy ExpansionSahil Mishra, Srinitish Srinivasan, Srikanta Bedathur, Tanmoy ChakrabortyWWW 2026 · 被引用 1 次
- Gaussian Mixture Variational Autoencoder with Contrastive Learning for Multi-Label ClassificationJunwen Bai, Shufeng Kong, Carla P. GomesICML 2022 · 被引用 48 次
- A Geometric Approach to Personalized Recommendation with Set-Theoretic Constraints Using Box EmbeddingsShib Sankar Dasgupta, Michael Boratko, Andrew McCallumICML 2025
