Exploiting Semantic Relations for Glass Surface Detection
Jiaying Lin, Yuen Hei Yeung, Rynson W. H. Lau
Abstract
Glass surfaces are omnipresent in our daily lives and often go unnoticed by the majority of us. While humans are generally able to infer their locations and thus avoid collisions, it can be difficult for current object detection systems to handle them due to the transparent nature of glass surfaces. Previous methods approached the problem by extracting global context information to obtain priors such as object boundaries and reflections. However, their performances cannot be guaranteed when these deterministic features are not available. We observe that humans often reason through the semantic context of the environment, which offers insights into the categories of and proximity between entities that are expected to appear in the surrounding. For example, the odds of the co-occurrence of glass windows with walls and curtains are generally higher than that with other objects, such as cars and trees, which have relatively less semantic relevance. Based on this observation, we propose a model named Glass Semantic Network ('GlassSemNet') that integrates the contextual relationship of the scenes for glass surface detection with two novel modules: (1) Scene Aware Activation (SAA) Module to adaptively filter critical channels with respect to spatial and semantic features, and (2) Context Correlation Attention (CCA) Module to progressively learn the contextual correlations among objects both spatially and semantically. In addition, we propose a large-scale glass surface detection dataset named Glass Surface Detection -Semantics ('GSD-S'), which contains 4,519 real-world RGB glass surface images from diverse real-world scenes with detailed annotations for both glass surface detection and semantic segmentation. Experimental results show that our model outperforms state-of-the-art works, especially with 42.6% MAE improvement on our proposed GSD-S dataset. Code, dataset, and models are available at https: // jiaying . link/ neurips2022-gsds/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5b9b2832-287d-4f6e-8d6a-2b980df0b671Cited by top-tier papers7
- Leveraging RGB-D Data with Cross-Modal Context Mining for Glass Surface DetectionJiaying Lin, Yuen Hei Yeung, Shuquan Ye, Rynson W. H. LauAAAI 2025 · 15 citations
- Multi-View Dynamic Reflection Prior for Video Glass Surface DetectionFang Liu, Yuhao Liu, Jiaying Lin, Ke Xu et al.AAAI 2024 · 12 citations
- GlassWizard: Harvesting Diffusion Priors for Glass Surface DetectionWenxue Li, Tian Ye, Xinyu Xiong, Jinbin Bai et al.ICCV 2025 · 8 citations
- Controllable-Lpmoe: Adapting to Challenging Object Segmentation Via Dynamic Local Priors From Mixture-Of-ExpertsYanguang Sun, Jiawei Lian, Jian Yang, Lei LuoICCV 2025 · 4 citations
- DepthFocus: Controllable Depth Estimation for See-Through Scenesjunhong min, Jimin Kim, Minwook Kim, Cheol-Hui Min et al.CVPR 2026 · 4 citations
Builds on20
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNetLi Yuan, Yunpeng Chen, Tao Wang, Weihao Yu et al.ICCV 2021 · 2,462 citations
- Twins: Revisiting the Design of Spatial Attention in Vision TransformersXiangxiang Chu, Zhi Tian, Yuqing Wang, Bo Zhang et al.NeurIPS 2021 · 1,388 citations
Related papers
- Multi-Semantic Modeling for Glass Surface Detection in the WildQianyu Cheng, Huankang Guan, Rynson W. H. LauAAAI 2026
- Don't Hit Me! Glass Detection in Real-World ScenesHaiyang Mei, Xin Yang, Yang Wang, Yuanyuan Liu et al.CVPR 2020
- Rich Context Aggregation With Reflection Prior for Glass Surface DetectionJiaying Lin, Zebang He, Rynson W. H. LauCVPR 2021
- MVGD-Net: A Novel Motion-aware Video Glass Surface Detection MethodYiwei Lu, Hao Huang, Tao YanAAAI 2026
- Scene Context-Aware Salient Object DetectionAvishek Siris, Jianbo Jiao, Gary K. L. Tam, Xianghua Xie et al.ICCV 2021 · 54 citations
