Tree Instance Segmentation with Temporal Contour Graph
Adnan Firoze, Cameron Wingren, Raymond A. Yeh, Bedrich Benes, Daniel G. Aliaga
Abstract
We present a novel approach to perform instance segmentation and counting for densely packed self-similar trees using a top-view RGB image sequence. We propose a solution that leverages pixel content, shape, and self-occlusion. First, we perform an initial over-segmentation of the image sequence and aggregate structural characteristics into a contour graph with temporal information incorporated. Second, using a graph convolutional network and its inherent local messaging passing abilities, we merge adjacent tree crown patches into a final set of tree crowns. Per various studies and comparisons, our method is superior to all prior methods and results in high-accuracy instance segmentation and counting despite the trees being tightly packed. Finally, we provide various forest image sequence datasets suitable for subsequent benchmarking and evaluation captured at different altitudes and leaf conditions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- GlobalMapper: Arbitrary-Shaped Urban Layout GenerationLiu He, Daniel G. AliagaICCV 2023 · 22 citations
- Bringing SAM to new heights: leveraging elevation data for tree crown segmentation from drone imageryMélisande Teng, Arthur Ouaknine, Etienne Laliberté, Yoshua Bengio et al.NeurIPS 2025 · 9 citations
- SelvaBox: A high‑resolution dataset for tropical tree crown detectionHugo Baudchon, Arthur Ouaknine, Martin Weiss, Mélisande Teng et al.ICLR 2026 · 8 citations
- SVDTree: Semantic Voxel Diffusion for Single Image Tree ReconstructionYuan Li, Zhihao Liu, Bedrich Benes, Xiaopeng Zhang et al.CVPR 2024
- Neural Hierarchical Decomposition for Single Image Plant ModelingZhihao Liu, Zhanglin Cheng, Naoto YokoyaCVPR 2025
Builds on13
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- Swin Transformer V2: Scaling Up Capacity and ResolutionZe Liu, Han Hu, Yutong Lin, Zhuliang Yao et al.CVPR 2022 · 2,138 citations
- YOLACT: Real-Time Instance SegmentationDaniel Bolya, Chong Zhou, Fanyi Xiao, Yong Jae LeeICCV 2019 · 2,075 citations
- Pixel Difference Networks for Efficient Edge DetectionZhuo Su, Wenzhe Liu, Zitong Yu, Dewen Hu et al.ICCV 2021 · 488 citations
Related papers
- Learning to Model Pixel-Embedded Affinity for Homogeneous Instance SegmentationWei Huang, Shiyu Deng, Chang Chen, Xueyang Fu et al.AAAI 2022 · 27 citations
- SelfSAGCN: Self-Supervised Semantic Alignment for Graph Convolution NetworkXu Yang, Cheng Deng, Zhiyuan Dang, Kun Wei et al.CVPR 2021
- Representation Learning of Geometric TreesZheng Zhang, Allen Zhang, Ruth Nelson, Giorgio A. Ascoli et al.KDD 2024
- Space-Time-Separable Graph Convolutional Network for Pose ForecastingTheodoros Sofianos, Alessio Sampieri, Luca Franco, Fabio GalassoICCV 2021 · 188 citations
- Self-Supervised Visual Representation Learning from Hierarchical GroupingXiao Zhang, Michael MaireNeurIPS 2020 · 82 citations
