Video-based Transparent Object Segmentation via Temporal Feature Aggregation
Zhen Wang, Dongyuan Li, Yaozu Wu, Peide Zhu, Shiyin Tan, Renhe Jiang
Abstract
Transparent object segmentation from a single image has been investigated for several years. However, detecting transparent areas from video has not been well explored, especially for different kinds of transparent categories besides glass, due to the scarcity of such a dataset. Therefore, in this paper, we propose the video-based transparent object segmentation task and introduce the first-of-its-kind corresponding dataset named TransVid, which contains nearly 400 videos with a total of 18,523 frames. Based on TranVid, we further propose a new method called TranSeg, in which we innovatively introduce Graph Neural Networks into the temporal segmentation task and combined with a novel Diffusion Model to make the model's segmentation results more accurate. Experimental results show that TranSeg achieves higher accuracy with fewer parameters than previous state-of-the-art models, demonstrating the effectiveness of our method. Moreover, comprehensive ablation analysis reveal several fascinating insights and suggest viable paths for further research.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get fb889f64-c87e-448d-aaee-98550bc0bbe6Related papers
- Enhanced Boundary Learning for Glass-like Object SegmentationHao He, Xiangtai Li, Guangliang Cheng, Jianping Shi et al.ICCV 2021 · 107 citations
- Transparent Object Tracking BenchmarkHeng Fan, Halady Akhilesha Miththanthaya, Harshit, Siranjiv Ramana Rajan et al.ICCV 2021 · 32 citations
- Multi-View Dynamic Reflection Prior for Video Glass Surface DetectionFang Liu, Yuhao Liu, Jiaying Lin, Ke Xu et al.AAAI 2024 · 12 citations
- Multi-view Spectral Polarization Propagation for Video Glass SegmentationYu Qiao, Bo Dong, Ao Jin, Yu Fu et al.ICCV 2023 · 9 citations
- 2D Gaussian Splatting-Based Sparse-View Transparent Object Depth Reconstruction Via Physics Simulation for Scene UpdateJeongyun Kim, Seunghoon Jeong, Giseop Kim, Myung-Hwan Jeon et al.ICCV 2025 · 1 citation
