Do It Yourself: Learning Semantic Correspondence from Pseudo-Labels
Olaf Dünkel, Thomas Wimmer, Christian Theobalt, Christian Rupprecht, Adam Kortylewski
Abstract
Finding correspondences between semantically similar points across images and object instances is one of the everlasting challenges in computer vision. While large pretrained vision models have recently been demonstrated as effective priors for semantic matching, they still suffer from ambiguities for symmetric objects or repeated object parts. We propose improving semantic correspondence estimation through 3D-aware pseudo-labeling. Specifically, we train an adapter to refine off-the-shelf features using pseudolabels obtained via 3D-aware chaining, filtering wrong labels through relaxed cyclic consistency, and 3D spherical prototype mapping constraints. While reducing the need for dataset-specific annotations compared to prior work, we establish a new state-of-the-art on SPair-71k, achieving an absolute gain of over 4% and of over 7% compared to methods with similar supervision requirements. The generality of our proposed approach simplifies the extension of training to other data sources, which we demonstrate in our experiments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 05884683-eea7-4f8b-b1e0-051b9aa1e84cCited by top-tier papers4
- Attention (as Discrete-Time Markov) ChainsYotam Erel, Olaf Dünkel, Rishabh Dabral, Vladislav Golyanik et al.NeurIPS 2025 · 12 citations
- MARCO: Navigating the Unseen Space of Semantic CorrespondenceClaudia Cuttano, Gabriele Trivigno, Carlo Masone, Stefan RothCVPR 2026 · 4 citations
- GECO: Geometrically Consistent Embedding with Lightspeed InferenceRegine Hartwig, Dominik Muhle, Riccardo Marin, Daniel CremersICCV 2025 · 2 citations
- Shape-of-You: Fused Gromov-Wasserstein Optimal Transport for Semantic Correspondence in-the-WildJiin Im, Sisung Liu, Je Hyeong HongCVPR 2026
Builds on32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Self-labelling via simultaneous clustering and representation learningYuki Markus Asano, Christian Rupprecht, Andrea VedaldiICLR 2020 · 873 citations
- A Tale of Two Features: Stable Diffusion Complements DINO for Zero-Shot Semantic CorrespondenceJunyi Zhang, Charles Herrmann, Junhwa Hur, Luisa Polania Cabrera et al.NeurIPS 2023 · 371 citations
Related papers
- Improving Semantic Correspondence with Viewpoint-Guided Spherical MapsOctave Mariotti, Oisin Mac Aodha, Hakan BilenCVPR 2024 · 12 citations
- Telling Left from Right: Identifying Geometry-Aware Semantic CorrespondenceJunyi Zhang, Charles Herrmann, Junhwa Hur, Eric Chen et al.CVPR 2024
- SemAlign3D: Semantic Correspondence between RGB-Images through Aligning 3D Object-Class RepresentationsKrispin Wandel, Hesheng WangCVPR 2025
- Bridging Viewpoint Gaps: Geometric Reasoning Boosts Semantic CorrespondenceQiyang Qian, Hansheng Chen, Masayoshi Tomizuka, Kurt Keutzer et al.CVPR 2025
- Semi-Supervised Learning of Semantic Correspondence with Pseudo-LabelsJiwon Kim, Kwangrok Ryoo, Junyoung Seo, Gyuseong Lee et al.CVPR 2022 · 23 citations
