RnGCam: High-Speed Video from Rolling & Global Shutter Measurements
Kevin Tandi, Xiang Dai, Chinmay Talegaonkar, Gal Mishne, Nick Antipa
Abstract
Compressive video capture encodes a short high-speed video into a single measurement using a low-speed sensor, then computationally reconstructs the original video. Prior implementations rely on expensive hardware and are restricted to imaging sparse scenes with empty backgrounds. We propose RnGCam, a system that fuses measurements from low-speed consumer-grade rolling-shutter (RS) and global-shutter (GS) sensors into video at kHz frame rates. The RS sensor is combined with a pseudorandom optic, called a diffuser, which spatially multiplexes scene information. The GS sensor is coupled with a conventional lens. The RS-diffuser provides low spatial detail and high temporal detail, complementing the GS-lens system's high spatial detail and low temporal detail. We propose a reconstruction method using implicit neural representations (INR) to fuse the measurements into a high-speed video. Our INR method separately models the static and dynamic scene components, while explicitly regularizing dynamics. In simulation, we show that our approach significantly outperforms previous RS compressive video methods, as well as state-of-the-art frame interpolators. We validate our approach in a dual-camera hardware setup, which generates 230 frames of video at 4,800 frames per second for dense scenes, using hardware that costs less than previous compressive video systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on14
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- Video Frame Interpolation with TransformerLiying Lu, Ruizheng Wu, Huaijia Lin, Jiangbo Lu et al.CVPR 2022 · 128 citations
- Time Lens++: Event-based Frame Interpolation with Parametric Nonlinear Flow and Multi-scale FusionStepan Tulyakov, Alfredo Bochicchio, Daniel Gehrig, Stamatios Georgoulis et al.CVPR 2022 · 126 citations
- Bacon: Band-limited Coordinate Networks for Multiscale Scene RepresentationDavid B. Lindell, Dave Van Veen, Jeong Joon Park, Gordon WetzsteinCVPR 2022 · 105 citations
Related papers
- Event-guided Frame Interpolation and Dynamic Range Expansion of Single Rolling Shutter ImageGuixu Lin, Jin Han, Mingdeng Cao, Zhihang Zhong et al.ACM MM 2023 · 11 citations
- SCINeRF: Neural Radiance Fields from a Snapshot Compressive ImageYunhao Li, Xiaodong Wang, Ping Wang, Xin Yuan et al.CVPR 2024
- USB-NeRF: Unrolling Shutter Bundle Adjusted Neural Radiance FieldsMoyang Li, Peng Wang, Lingzhe Zhao, Bangyan Liao et al.ICLR 2024 · 13 citations
- RECOMBINER: Robust and Enhanced Compression with Bayesian Implicit Neural RepresentationsJiajun He, Gergely Flamich, Zongyu Guo, José Miguel Hernández-LobatoICLR 2024 · 12 citations
- Towards HDR and HFR Video from Rolling-Mixed-Bit SpikingsYakun Chang, Yeliduosi Xiaokaiti, Yujia Liu, Bin Fan et al.CVPR 2024
