Spatial-Temporal Super-Resolution of Satellite Imagery via Conditional Pixel Synthesis
Yutong He, Dingjie Wang, Nicholas Lai, William Zhang, Chenlin Meng, Marshall Burke, David B. Lobell, Stefano Ermon
Abstract
High-resolution satellite imagery has proven useful for a broad range of tasks, including measurement of global human population, local economic livelihoods, and biodiversity, among many others. Unfortunately, high-resolution imagery is both infrequently collected and expensive to purchase, making it hard to efficiently and effectively scale these downstream tasks over both time and space. We propose a new conditional pixel synthesis model that uses abundant, low-cost, low-resolution imagery to generate accurate high-resolution imagery at locations and times in which it is unavailable. We show that our model attains photo-realistic sample quality and outperforms competing baselines on a key downstream task -- object counting -- particularly in geographic locations where conditions on the ground are changing rapidly.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- SatMAE: Pre-training Transformers for Temporal and Multi-Spectral Satellite ImageryYezhen Cong, Samar Khanna, Chenlin Meng, Patrick Liu et al.NeurIPS 2022 · 707 citations
- Scale-MAE: A Scale-Aware Masked Autoencoder for Multiscale Geospatial Representation LearningColorado J. Reed, Ritwik Gupta, Shufan Li, Sarah Brockman et al.ICCV 2023 · 373 citations
- DiffusionSat: A Generative Foundation Model for Satellite ImagerySamar Khanna, Patrick Liu, Linqi Zhou, Chenlin Meng et al.ICLR 2024 · 173 citations
- CSP: Self-Supervised Contrastive Spatial Pre-Training for Geospatial-Visual RepresentationsGengchen Mai, Ni Lao, Yutong He, Jiaming Song et al.ICML 2023 · 103 citations
- 4KAgent: Agentic Any Image to 4K Super-ResolutionYushen Zuo, Qi Zheng, Mingyang Wu, Xinrui Jiang et al.NeurIPS 2025 · 51 citations
Builds on6
- GRAF: Generative Radiance Fields for 3D-Aware Image SynthesisKatja Schwarz, Yiyi Liao, Michael Niemeyer, Andreas GeigerNeurIPS 2020 · 1,001 citations
- Image Generators With Conditionally-Independent Pixel SynthesisIvan Anokhin, Kirill Demochkin, Taras Khakhulin, Gleb Sterkin et al.CVPR 2021
- pixelNeRF: Neural Radiance Fields From One or Few ImagesAlex Yu, Vickie Ye, Matthew Tancik, Angjoo KanazawaCVPR 2021
- Space-Time Neural Irradiance Fields for Free-Viewpoint VideoWenqi Xian, Jia-Bin Huang, Johannes Kopf, Changil KimCVPR 2021
- Neural Scene Flow Fields for Space-Time View Synthesis of Dynamic ScenesZhengqi Li, Simon Niklaus, Noah Snavely, Oliver WangCVPR 2021
Related papers
- Efficient Poverty Mapping from High Resolution Remote Sensing ImagesKumar Ayush, Burak Uzkent, Kumar Tanmay, Marshall Burke et al.AAAI 2021 · 51 citations
- Coming Down to Earth: Satellite-to-Street View Synthesis for Geo-LocalizationAysim Toker, Qunjie Zhou, Maxim Maximov, Laura Leal-TaixéCVPR 2021
- Cross-View Splatter: Feed-Forward View Synthesis with Georeferenced ImagesMatias Turkulainen, Akshay Krishnan, Filippo Aleotti, Mohamed Sayed et al.CVPR 2026
- PixelTransformer: Sample Conditioned Signal GenerationShubham Tulsiani, Abhinav GuptaICML 2021 · 18 citations
- Learning When and Where to Zoom With Deep Reinforcement LearningBurak Uzkent, Stefano ErmonCVPR 2020
