Object-level Geometric Structure Preserving for Natural Image Stitching
Wenxiao Cai, Wankou Yang
Abstract
The topic of stitching images with globally natural structures holds paramount significance, with two main goals: pixel-level alignment and distortion prevention. The existing approaches exhibit the ability to align well, yet fall short in maintaining object structures. In this paper, we endeavour to safeguard the overall OBJect-level structures within images based on Global Similarity Prior (OBJ-GSP), on the basis of good alignment performance. Our approach leverages semantic segmentation models like the family of Segment Anything Model to extract the contours of any objects in a scene. Triangular meshes are employed in image transformation to protect the overall shapes of objects within images. The balance between alignment and distortion prevention is achieved by allowing the object meshes to strike a balance between similarity and projective transformation. We also demonstrate that object-level semantic information is necessary in low-altitude aerial image stitching. Additionally, we propose StitchBench, the largest image stitching benchmark with most diverse scenarios. Extensive experimental results demonstrate that OBJ-GSP outperforms existing methods in both pixel alignment and shape preservation. Code and dataset is publicly available at https://github.com/RussRobin/OBJ-GSP .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e7a44ff6-999f-4543-bb36-20e49680bfc3Cited by top-tier papers2
- PixelStitch: Structure-Preserving Pixel-Wise Bidirectional Warps for Unsupervised Image StitchingHengzhe Jin, Lang Nie, Chunyu Lin, Xiaomei Feng et al.ICCV 2025 · 4 citations
- Towards Generalized Multimodal Homography EstimationJinkun You, Jiaxin Cheng, Jie Zhang, Yicong ZhouCVPR 2026 · 1 citation
Builds on5
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- EfficientSAM: Leveraged Masked Image Pretraining for Efficient Segment AnythingYunyang Xiong, Bala Varadarajan, Lemeng Wu, Xiaoyu Xiang et al.CVPR 2024 · 185 citations
- Parallax-Tolerant Unsupervised Deep Image StitchingLang Nie, Chunyu Lin, Kang Liao, Shuaicheng Liu et al.ICCV 2023 · 111 citations
- Geometric Structure Preserving Warp for Natural Image StitchingPeng Du, Jifeng Ning, Jiguang Cui, Shaoli Huang et al.CVPR 2022 · 56 citations
- Leveraging Line-Point Consistence To Preserve Structures for Wide Parallax Image StitchingQi Jia, Zhengjun Li, Xin Fan, Haotian Zhao et al.CVPR 2021
Related papers
- Learning Pixel-wise Alignment for Unsupervised Image StitchingQi Jia, Xiaomei Feng, Yu Liu, Xin Fan et al.ACM MM 2023 · 34 citations
- Preserve Anything: Controllable Image Synthesis with Object PreservationPrasen Kumar Sharma, Neeraj Matiyali, Siddharth Srivastava, Gaurav SharmaICCV 2025
- High Fidelity Aggregated Planar Prior Assisted PatchMatch Multi-View StereoJie Liang, Rongjie Wang, Rui Peng, Zhe Zhang et al.ACM MM 2024 · 3 citations
- Pixel-Wise Warping for Deep Image StitchingHyeokjun Kweon, Hyeonseong Kim, Yoonsu Kang, Youngho Yoon et al.AAAI 2023 · 24 citations
- Scene Grounding in the WildTamir Cohen, Leo Segre, Shay Shomer Chai, Shai Avidan et al.CVPR 2026 · 1 citation
