NeuriCam: Key-Frame Video Super-Resolution and Colorization for IoT Cameras
Bandhav Veluri, Collin Pernu, Ali Saffari, Joshua R. Smith, Michael B. Taylor, Shyamnath Gollakota
Abstract
We present NeuriCam, a novel deep learning-based system to achieve video capture from low-power dual-mode IoT camera systems. Our idea is to design a dual-mode camera system where the first mode is low power (1.1 mW) but only outputs grey-scale, low resolution and noisy video and the second mode consumes much higher power (100 mW) but outputs color and higher resolution images. To reduce total energy consumption, we heavily duty cycle the high power mode to output an image only once every second. The data for this camera system is then wirelessly sent to a nearby plugged-in gateway, where we run our real-time neural network decoder to reconstruct a higher-resolution color video. To achieve this, we introduce an attention feature filter mechanism that assigns different weights to different features, based on the correlation between the feature map and the contents of the input frame at each spatial location. We design a wireless hardware prototype using off-the-shelf cameras and address practical issues including packet loss and perspective mismatch. Our evaluations show that our dual-camera approach reduces energy consumption by 7x compared to existing systems. Further, our model achieves an average greyscale PSNR gain of 3.7 dB over prior single and dual-camera video super-resolution methods and 5.6 dB RGB gain over prior color propagation methods.
Open-source code: https://github.com/vb000/NeuriCam.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- IRIS: Wireless ring for vision-based smart home interactionMaruchi Kim, Antonio Glenn, Bandhav Veluri, Yunseo Lee et al.UIST 2024 · 14 citations
- SeaScan: An Energy-Efficient Underwater Camera for Wireless 3D Color ImagingNazish Naeem, Jack Rademacher, Ritik Patnaik, Tara Boroushaki et al.MobiCom 2024 · 7 citations
- HyperCam: Low-Power Onboard Computer Vision for IoT CamerasChae Young Lee, Pu (Luke) Yi, Maxwell Fite, Tejus Rao et al.MobiCom 2025 · 5 citations
- VueBuds: Visual Intelligence with Wireless EarbudsMaruchi Kim, Rasya Fawwaz, Zhi Yang Lim, Brinda Moudgalya et al.CHI 2026 · 1 citation
Builds on13
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision TransformerSachin Mehta, Mohammad RastegariICLR 2022 · 2,162 citations
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- Progressive Fusion Video Super-Resolution Network via Exploiting Non-Local Spatio-Temporal CorrelationsPeng Yi, Zhongyuan Wang, Kui Jiang, Junjun Jiang et al.ICCV 2019 · 309 citations
Related papers
- Neuromorphic Camera Guided High Dynamic Range ImagingJin Han, Chu Zhou, Peiqi Duan, Yehui Tang et al.CVPR 2020
- NeuroScaler: neural video enhancement at scaleHyunho Yeo, Hwijoon Lim, Jaehong Kim, Youngmok Jung et al.SIGCOMM 2022 · 54 citations
- Event Stream Super-Resolution via Spatiotemporal Constraint LearningSiqi Li, Yutong Feng, Yipeng Li, Yu Jiang et al.ICCV 2021 · 25 citations
- Event-based Motion Deblurring with Modality-Aware Decomposition and RecompositionWen Yang, Jinjian Wu, Leida Li, Weisheng Dong et al.ACM MM 2023 · 14 citations
- Learning for Motion Deblurring with Hybrid Frames and EventsWen Yang, Jinjian Wu, Jupo Ma, Leida Li et al.ACM MM 2022 · 17 citations
