TransiT: Transient Transformer for Non-Line-of-Sight Videography
Ruiqian Li, Siyuan Shen, Suan Xia, Ziheng Wang, Xingyue Peng, Chengxuan Song, Yingsheng Zhu, Tao Wu, Shiying Li, Jingyi Yu
摘要
High quality and high speed videography using Non-Line-of-Sight (NLOS) imaging benefit autonomous navigation, collision prevention, and post-disaster search and rescue tasks. Current solutions have to balance between the frame rate and image quality. High frame rates, for example, can be achieved by reducing either per-point scanning time or scanning density, but at the cost of lowering the information density at individual frames. Fast scanning process further reduces the signal-to-noise ratio and different scanning systems exhibit different distortion characteristics. In this work, we design and employ a new Transient Transformer architecture called TransiT to achieve real-time NLOS recovery under fast scans. TransiT directly compresses the temporal dimension of input transients to extract features, reducing computation costs and meeting high frame rate requirements. It further adopts a feature fusion mechanism as well as employs a spatial-temporal Transformer to help capture features of NLOS transient videos. Moreover, TransiT applies transfer learning to bridge the gap between synthetic and real-measured data. In real experiments, TransiT manages to reconstruct from sparse transients of measured at an exposure time of 0.4 ms per point to NLOS videos at a resolution at 10 frames per second. We will make our code and dataset available to the community.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper10
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra 等NeurIPS 2022 · 被引用 5,493 次
- ViViT: A Video Vision TransformerAnurag Arnab, Mostafa Dehghani, Georg Heigold, Chen Sun 等ICCV 2021 · 被引用 2,947 次
- Is Space-Time Attention All You Need for Video Understanding?Gedas Bertasius, Heng Wang, Lorenzo TorresaniICML 2021 · 被引用 2,927 次
- Video Swin TransformerZe Liu, Jia Ning, Yue Cao, Yixuan Wei 等CVPR 2022 · 被引用 1,847 次
- Deep Non-line-of-sight Imaging from Under-scanning MeasurementsYue Li, Yueyi Zhang, Juntian Ye, Feihu Xu 等NeurIPS 2023 · 被引用 32 次
相关 Paper
- Non-Line-of-Sight Imaging with Signal Superresolution NetworkJianyu Wang, Xintong Liu, Leping Xiao, Zuoqiang Shi 等CVPR 2023
- Virtual Scanning: Unsupervised Non-line-of-sight Imaging from Irregularly Undersampled TransientsXingyu Cui, Huanjing Yue, Song Li, Xiangjun Yin 等NeurIPS 2024 · 被引用 13 次
- NLOST: Non-Line-of-Sight Imaging with TransformerYue Li, Jiayong Peng, Juntian Ye, Yueyi Zhang 等CVPR 2023
- Toward Dynamic Non-Line-of-Sight Imaging with Mamba Enforced Temporal ConsistencyYue Li, Yi Sun, Shida Sun, Juntian Ye 等NeurIPS 2024 · 被引用 9 次
- EfficientSCI: Densely Connected Network with Space-time Factorization for Large-scale Video Snapshot Compressive ImagingLishun Wang, Miao Cao, Xin YuanCVPR 2023
