Video-based Transparent Object Segmentation via Temporal Feature Aggregation
Zhen Wang, Dongyuan Li, Yaozu Wu, Peide Zhu, Shiyin Tan, Renhe Jiang
摘要
Transparent object segmentation from a single image has been investigated for several years. However, detecting transparent areas from video has not been well explored, especially for different kinds of transparent categories besides glass, due to the scarcity of such a dataset. Therefore, in this paper, we propose the video-based transparent object segmentation task and introduce the first-of-its-kind corresponding dataset named TransVid, which contains nearly 400 videos with a total of 18,523 frames. Based on TranVid, we further propose a new method called TranSeg, in which we innovatively introduce Graph Neural Networks into the temporal segmentation task and combined with a novel Diffusion Model to make the model's segmentation results more accurate. Experimental results show that TranSeg achieves higher accuracy with fewer parameters than previous state-of-the-art models, demonstrating the effectiveness of our method. Moreover, comprehensive ablation analysis reveal several fascinating insights and suggest viable paths for further research.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Enhanced Boundary Learning for Glass-like Object SegmentationHao He, Xiangtai Li, Guangliang Cheng, Jianping Shi 等ICCV 2021 · 被引用 107 次
- Transparent Object Tracking BenchmarkHeng Fan, Halady Akhilesha Miththanthaya, Harshit, Siranjiv Ramana Rajan 等ICCV 2021 · 被引用 32 次
- Multi-View Dynamic Reflection Prior for Video Glass Surface DetectionFang Liu, Yuhao Liu, Jiaying Lin, Ke Xu 等AAAI 2024 · 被引用 12 次
- Multi-view Spectral Polarization Propagation for Video Glass SegmentationYu Qiao, Bo Dong, Ao Jin, Yu Fu 等ICCV 2023 · 被引用 9 次
- 2D Gaussian Splatting-Based Sparse-View Transparent Object Depth Reconstruction Via Physics Simulation for Scene UpdateJeongyun Kim, Seunghoon Jeong, Giseop Kim, Myung-Hwan Jeon 等ICCV 2025 · 被引用 1 次
