Improving Semantic Correspondence with Viewpoint-Guided Spherical Maps
Octave Mariotti, Oisin Mac Aodha, Hakan Bilen
Abstract
Recent self-supervised models produce visual features that are not only effective at encoding image-level, but also pixel-level, semantics. They have been reported to obtain impressive results for dense visual semantic correspondence estimation, even outperforming fully-supervised methods. Nevertheless, these models still fail in the pres-ence of challenging image characteristics such as symme-tries and repeated parts. To address these limitations, we propose a new semantic correspondence estimation method that supplements state-of-the-art self-supervised features with 3D understanding via a weak geometric spherical prior. Compared to more involved 3D pipelines, our model provides a simple and effective way of injecting informative geometric priors into the learned representation while requiring only weak viewpoint information. We also propose a new evaluation metric that better accounts for re-peated part and symmetry-induced mistakes. We show that our method succeeds in distinguishing between symmetric views and repeated parts across many object categories in the challenging SPair-71 k dataset and also in generalizing to previously unseen classes in the AwA dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e8b15563-ae58-4f9b-92b4-11c7cac3ae11Cited by top-tier papers18
- Jamais Vu: Exposing the Generalization Gap in Supervised Semantic CorrespondenceOctave Mariotti, Zhipeng Du, Yash Bhalgat, Oisin Mac Aodha et al.NeurIPS 2025 · 8 citations
- Learning 3D Object Spatial Relationships From Pre-Trained 2D Diffusion ModelsSangwon Baik, Hyeonwoo Kim, Hanbyul JooICCV 2025 · 7 citations
- Do It Yourself: Learning Semantic Correspondence from Pseudo-LabelsOlaf Dünkel, Thomas Wimmer, Christian Theobalt, Christian Rupprecht et al.ICCV 2025 · 4 citations
- MARCO: Navigating the Unseen Space of Semantic CorrespondenceClaudia Cuttano, Gabriele Trivigno, Carlo Masone, Stefan RothCVPR 2026 · 4 citations
- GECO: Geometrically Consistent Embedding with Lightspeed InferenceRegine Hartwig, Dominik Muhle, Riccardo Marin, Daniel CremersICCV 2025 · 2 citations
Builds on19
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- A Tale of Two Features: Stable Diffusion Complements DINO for Zero-Shot Semantic CorrespondenceJunyi Zhang, Charles Herrmann, Junhwa Hur, Luisa Polania Cabrera et al.NeurIPS 2023 · 371 citations
- Unsupervised Semantic Correspondence Using Stable DiffusionEric Hedlin, Gopal Sharma, Shweta Mahajan, Hossam Isack et al.NeurIPS 2023 · 152 citations
Related papers
- Telling Left from Right: Identifying Geometry-Aware Semantic CorrespondenceJunyi Zhang, Charles Herrmann, Junhwa Hur, Eric Chen et al.CVPR 2024
- SemAlign3D: Semantic Correspondence between RGB-Images through Aligning 3D Object-Class RepresentationsKrispin Wandel, Hesheng WangCVPR 2025
- Bridging Viewpoint Gaps: Geometric Reasoning Boosts Semantic CorrespondenceQiyang Qian, Hansheng Chen, Masayoshi Tomizuka, Kurt Keutzer et al.CVPR 2025
- Semi-Supervised Learning of Semantic Correspondence with Pseudo-LabelsJiwon Kim, Kwangrok Ryoo, Junyoung Seo, Gyuseong Lee et al.CVPR 2022 · 23 citations
- ViewNet: Unsupervised Viewpoint Estimation from Conditional GenerationOctave Mariotti, Oisin Mac Aodha, Hakan BilenICCV 2021 · 8 citations
