Panoptic Segmentation of Satellite Image Time Series with Convolutional Temporal Attention Networks
Vivien Sainte Fare Garnot, Loïc Landrieu
Abstract
Unprecedented access to multi-temporal satellite imagery has opened new perspectives for a variety of Earth observation tasks. Among them, pixel-precise panoptic segmentation of agricultural parcels has major economic and environmental implications. While researchers have explored this problem for single images, we argue that the complex temporal patterns of crop phenology are better addressed with temporal sequences of images. In this paper, we present the first end-to-end, single-stage method for panoptic segmentation of Satellite Image Time Series (SITS). This module can be combined with our novel image sequence encoding network which relies on temporal self-attention to extract rich and adaptive multi-scale spatiotemporal features. We also introduce PASTIS, the first open-access SITS dataset with panoptic annotations. We demonstrate the superiority of our encoder for semantic segmentation against multiple competing architectures, and set up the first state-of-the-art of panoptic segmentation of SITS. Our implementation and PASTIS are publicly available.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers16
- SatMAE: Pre-training Transformers for Temporal and Multi-Spectral Satellite ImageryYezhen Cong, Samar Khanna, Chenlin Meng, Patrick Liu et al.NeurIPS 2022 · 707 citations
- SatlasPretrain: A Large-Scale Dataset for Remote Sensing Image UnderstandingFavyen Bastani, Piper Wolters, Ritwik Gupta, Joe Ferdinando et al.ICCV 2023 · 216 citations
- DynamicEarthNet: Daily Multi-Spectral Satellite Dataset for Semantic Change SegmentationAysim Toker, Lukas Kondmann, Mark Weber, Marvin Eisenberger et al.CVPR 2022 · 108 citations
- OlmoEarth: Stable Latent Image Modeling for Multimodal Earth ObservationHenry Herzog, Favyen Bastani, Yawen Zhang, Gabriel Tseng et al.CVPR 2026 · 28 citations
- GEOBench-VLM: Benchmarking Vision-Language Models for Geospatial TasksMuhammad Sohail Danish, Muhammad Akhtar Munir, Syed Roshaan Ali Shah, Kartik Kuckreja et al.ICCV 2025 · 11 citations
Builds on3
- Video Instance SegmentationLinjie Yang, Yuchen Fan, Ning XuICCV 2019 · 615 citations
- CenterMask: Single Shot Instance Segmentation With Point RepresentationYuqing Wang, Zhaoliang Xu, Hao Shen, Baoshan Cheng et al.CVPR 2020
- Satellite Image Time Series Classification With Pixel-Set Encoders and Temporal Self-AttentionVivien Sainte Fare Garnot, Loïc Landrieu, Sébastien Giordano, Nesrine ChehataCVPR 2020
Related papers
- ViTs for SITS: Vision Transformers for Satellite Image Time SeriesMichail Tarasiou, Erik Chavez, Stefanos ZafeiriouCVPR 2023
- Exact: Exploring Space-Time Perceptive Clues for Weakly Supervised Satellite Image Time Series Semantic SegmentationHao Zhu, Yan Zhu, Jiayu Xiao, Tianxiang Xiao et al.CVPR 2025
- Open-vocabulary Panoptic Segmentation with Embedding ModulationXi Chen, Shuang Li, Ser-Nam Lim, Antonio Torralba et al.ICCV 2023 · 42 citations
- EOV-Seg: Efficient Open-Vocabulary Panoptic SegmentationHongwei Niu, Jie Hu, Jianghang Lin, Guannan Jiang et al.AAAI 2025 · 11 citations
- Real-Time Panoptic Segmentation From Dense DetectionsRui Hou, Jie Li, Arjun Bhargava, Allan Raventos et al.CVPR 2020
