BF-STVSR: B-Splines and Fourier - Best Friends for High Fidelity Spatial-Temporal Video Super-Resolution
Eunjin Kim, Hyeonjin Kim, Kyong Hwan Jin, Jaejun Yoo
Abstract
While prior methods in Continuous Spatial-Temporal Video Super-Resolution (C-STVSR) employ Implicit Neural Representation (INR) for continuous encoding, they often struggle to capture the complexity of video data, relying on simple coordinate concatenation and pre-trained optical flow networks for motion representation. Interestingly, we find that adding position encoding, contrary to common observations, does not improve-and even degradesperformance. This issue becomes particularly pronounced when combined with pre-trained optical flow networks, which can limit the model's flexibility. To address these issues, we propose BF-STVSR, a C-STVSR framework with two key modules tailored to better represent spatial and temporal characteristics of video: 1) B-spline Mapper for smooth temporal interpolation, and 2) Fourier Mapper for capturing dominant spatial frequencies. Our approach achieves state-of-the-art in various metrics, including PSNR and SSIM, showing enhanced spatial details and natural temporal consistency. Our code is available here.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Continuous Space-Time Video Super-Resolution with 3D Fourier FieldsAlexander Becker, Julius Erbach, Dominik Narnhofer, Konrad SchindlerICLR 2026 · 3 citations
- Spatio-Temporal Distortion Aware Omnidirectional Video Super-ResolutionHongyu An, Xinfeng Zhang, Shijie Zhao, Li Zhang et al.AAAI 2026 · 3 citations
- MambaOVSR: Multiscale Fusion with Global Motion Modeling for Chinese Opera Video Super-ResolutionHua Chang, Xin Xu, Wei Liu, Wei Wang et al.AAAI 2026 · 1 citation
- Time Without Time: Pseudo-Temporal Representation for Space-Time Super-ResolutionHee Min Choi, Hyoa Kang, Suji Kim, Dokwan Oh et al.CVPR 2026
- GeoFAR: Geography-Informed Frequency-Aware Super-Resolution for Climate DataChang Xu, Gencer Sumbul, Li Mi, Robin Zbinden et al.ICLR 2026
Builds on19
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- Local Texture Estimator for Implicit Representation FunctionJaewon Lee, Kyong Hwan JinCVPR 2022 · 193 citations
Related papers
- VideoINR: Learning Video Implicit Neural Representation for Continuous Space-Time Super-ResolutionZeyuan Chen, Yinbo Chen, Jingwen Liu, Xingqian Xu et al.CVPR 2022 · 95 citations
- MoTIF: Learning Motion Trajectories with Local Implicit Neural Functions for Continuous Space-Time Video Super-ResolutionYi-Hsin Chen, Si-Cun Chen, Yi-Hsin Chen, Yen-Yu Lin et al.ICCV 2023 · 27 citations
- Bias for Action: Video Implicit Neural Representations with Bias ModulationAlper Kayabasi, Anil Kumar Vadathya, Guha Balakrishnan, Vishwanath SaragadamCVPR 2025
- Look Back and Forth: Video Super-Resolution with Explicit Temporal Difference ModelingTakashi Isobe, Xu Jia, Xin Tao, Changlin Li et al.CVPR 2022 · 57 citations
- DS-NeRV: Implicit Neural Video Representation with Decomposed Static and Dynamic CodesHao Yan, Zhihui Ke, Xiaobo Zhou, Tie Qiu et al.CVPR 2024 · 18 citations
