LogicSeg: Parsing Visual Semantics with Neural Logic Learning and Reasoning
Liulei Li, Wenguan Wang, Yang Yi
Abstract
Current high-performance semantic segmentation models are purely data-driven sub-symbolic approaches and blind to the structured nature of the visual world. This is in stark contrast to human cognition which abstracts visual perceptions at multiple levels and conducts symbolic reasoning with such structured abstraction. To fill these fundamental gaps, we devise LogicSeg, a holistic visual semantic parser that integrates neural inductive learning and logic reasoning with both rich data and symbolic knowledge. In particular, the semantic concepts of interest are structured as a hierarchy, from which a set of constraints are derived for describing the symbolic relations and formalized as first-order logic rules. After fuzzy logic-based continuous relaxation, logical formulae are grounded onto data and neural computational graphs, hence enabling logic-induced network training. During inference, logical constraints are packaged into an iterative process and injected into the network in a form of several matrix multiplications, so as to achieve hierarchy-coherent prediction with logic reasoning. These designs together make LogicSeg a general and compact neural-logic machine that is readily integrated into existing segmentation models. Extensive experiments over four datasets with various segmentation models and backbones verify the effectiveness and generality of LogicSeg.We believe this study opens a new avenue for visual semantic parsing.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8e2335b5-524f-4c7b-ba30-129d0b5bf0ceCited by top-tier papers13
- Neural-Logic Human-Object Interaction DetectionLiulei Li, Jianan Wei, Wenguan Wang, Yi YangNeurIPS 2023 · 54 citations
- Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion ModelsLiulei Li, Wenguan Wang, Yi YangNeurIPS 2024 · 29 citations
- Transferring to Real-World Layouts: A Depth-aware Framework for Scene AdaptationMu Chen, Zhedong Zheng, Yi YangACM MM 2024 · 19 citations
- Discriminative Perception via Anchored Description for Reasoning SegmentationTao Yang, Qing Zhou, Yanliang Li, Qi WangCVPR 2026 · 4 citations
- Zero-shot Compositional Action Recognition with Neural Logic ConstraintsGefan Ye, Lin Li, Kexin Li, Jun Xiao et al.ACM MM 2025 · 1 citation
Builds on43
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- SegNeXt: Rethinking Convolutional Attention Design for Semantic SegmentationMeng-Hao Guo, Cheng-Ze Lu, Qibin Hou, Zhengning Liu et al.NeurIPS 2022 · 1,385 citations
Related papers
- Logic-induced Diagnostic Reasoning for Semi-supervised Semantic SegmentationChen Liang, Wenguan Wang, Jiaxu Miao, Yi YangICCV 2023 · 55 citations
- Generating by Understanding: Neural Visual Generation with Logical Symbol GroundingsYifei Peng, Zijie Zha, Yu Jin, Zhexu Luo et al.KDD 2025
- Exploit Visual Dependency Relations for Semantic SegmentationMingyuan Liu, Dan Schonfeld, Wei TangCVPR 2021
- Deep Hierarchical Semantic SegmentationLiulei Li, Tianfei Zhou, Wenguan Wang, Jianwu Li et al.CVPR 2022 · 181 citations
- When Logic Meets Perception: Operator-Agnostic Differentiable Reasoning for Reliable Neural PredictionZihan Shao, Chang Lu, Renate A. Schmidt, Yizheng ZhaoKDD 2026
