Toward Dynamic Non-Line-of-Sight Imaging with Mamba Enforced Temporal Consistency
Yue Li, Yi Sun, Shida Sun, Juntian Ye, Yueyi Zhang, Feihu Xu, Zhiwei Xiong
Abstract
Dynamic reconstruction in confocal non-line-of-sight imaging encounters great challenges since the dense raster-scanning manner limits the practical frame rate. A fewer pioneer works reconstruct high-resolution volumes from the under-scanning transient measurements but overlook temporal consistency among transient frames. To fully exploit multi-frame information, we propose the first spatial-temporal Mamba (ST-Mamba) based method tailored for dynamic reconstruction of transient videos. Our method capitalizes on neighbouring transient frames to aggregate the target 3D hidden volume. Specifically, the interleaved features extracted from the input transient frames are fed to the proposed ST-Mamba blocks, which leverage the time-resolving causality in transient measurement. The cross ST-Mamba blocks are then devised to integrate the adjacent transient features. The target high-resolution transient frame is subsequently recovered by the transient spreading module. After transient fusion and recovery, a physical-based network is employed to reconstruct the hidden volume. To tackle the substantial noise inherent in transient videos, we propose a wave-based loss function to impose constraints within the phasor field. Besides, we introduce a new dataset, comprising synthetic videos for training and real-world videos for evaluation. Extensive experiments showcase the superior performance of our method on both synthetic data and real-world data captured by different imaging setups. The code and data are available at https://github.com/Depth2World/Dynamic_NLOS .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 849cbe10-ec7a-43f7-8f80-af3c16ea3e24Cited by top-tier papers2
- DENALI: A Dataset Enabling Non-Line-of-Sight Spatial Reasoning with Low-Cost LiDARsNikhil Behari, Diego Rivero, Luke Apostolides, Suman Ghosh et al.CVPR 2026 · 2 citations
- Non-line-of-sight imaging with arbitrary relay surface geometries via 3D Gaussian Transient RenderingYi Wang, Ziyu Zhan, Yuran Wang, Hao Wang et al.SIGGRAPH 2026
Builds on12
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 3,482 citations
- VMamba: Visual State Space ModelYue Liu, Yunjie Tian, Yuzhong Zhao, Hongtian Yu et al.NeurIPS 2024 · 3,199 citations
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
- PointMamba: A Simple State Space Model for Point Cloud AnalysisDingkang Liang, Xin Zhou, Wei Xu, Xingkui Zhu et al.NeurIPS 2024 · 380 citations
- Deep Non-line-of-sight Imaging from Under-scanning MeasurementsYue Li, Yueyi Zhang, Juntian Ye, Feihu Xu et al.NeurIPS 2023 · 32 citations
Related papers
- NLOST: Non-Line-of-Sight Imaging with TransformerYue Li, Jiayong Peng, Juntian Ye, Yueyi Zhang et al.CVPR 2023
- TransiT: Transient Transformer for Non-Line-of-Sight VideographyRuiqian Li, Siyuan Shen, Suan Xia, Ziheng Wang et al.ICCV 2025 · 1 citation
- Virtual Scanning: Unsupervised Non-line-of-sight Imaging from Irregularly Undersampled TransientsXingyu Cui, Huanjing Yue, Song Li, Xiangjun Yin et al.NeurIPS 2024 · 13 citations
- EVDM: Event-based Real-World Video Deblurring with MambaZhijing Sun, Senyan Xu, Kean Liu, Runze Tian et al.ICCV 2025 · 6 citations
- Sp3ctralMamba: Physics-Driven Joint State Space Model for Hyperspectral Image ReconstructionGe Meng, Jingyan Tu, Jingjia Huang, Yunlong Lin et al.AAAI 2025 · 9 citations
