Unsupervised Learning of Landmarks by Descriptor Vector Exchange
James Thewlis, Samuel Albanie, Hakan Bilen, Andrea Vedaldi
Abstract
Equivariance to random image transformations is an effective method to learn landmarks of object categories, such as the eyes and the nose in faces, without manual supervision. However, this method does not explicitly guarantee that the learned landmarks are consistent with changes between different instances of the same object, such as different facial identities. In this paper, we develop a new perspective on the equivariance approach by noting that dense landmark detectors can be interpreted as local image descriptors equipped with invariance to intra-category variations. We then propose a direct method to enforce such an invariance in the standard equivariant loss. We do so by exchanging descriptor vectors between images of different object instances prior to matching them geometrically. In this manner, the same vectors must work regardless of the specific object identity considered. We use this approach to learn vectors that can simultaneously be interpreted as local descriptors and dense landmarks, combining the advantages of both. Experiments on standard benchmarks show that this approach can match, and in some cases surpass state-of-the-art performance amongst existing methods that learn landmarks without supervision. Code is available at www.robots.ox.ac.uk/ vgg/research/DVE/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 722e9217-cfba-4163-aadf-0566390e5e1cCited by top-tier papers32
- With a Little Help from My Friends: Nearest-Neighbor Contrastive Learning of Visual RepresentationsDebidatta Dwibedi, Yusuf Aytar, Jonathan Tompson, Pierre Sermanet et al.ICCV 2021 · 542 citations
- Unsupervised Learning of Dense Visual RepresentationsPedro O. Pinheiro, Amjad Almahairi, Ryan Y. Benmalek, Florian Golemo et al.NeurIPS 2020 · 227 citations
- Continuous Surface EmbeddingsNatalia Neverova, David Novotný, Marc Szafraniec, Vasil Khalidov et al.NeurIPS 2020 · 116 citations
- Unsupervised Object-Level Representation Learning from Scene ImagesJiahao Xie, Xiaohang Zhan, Ziwei Liu, Yew Soon Ong et al.NeurIPS 2021 · 93 citations
- Correspondence learning via linearly-invariant embeddingRiccardo Marin, Marie-Julie Rakotosaona, Simone Melzi, Maks OvsjanikovNeurIPS 2020 · 82 citations
Related papers
- Exploiting Invariance of Mining Facial LandmarksJiangming Shi, Zixian Gao, Hao Liu, Zekuan Yu et al.ACM MM 2021 · 2 citations
- On Equivariant and Invariant Learning of Object Landmark RepresentationsZezhou Cheng, Jong-Chyi Su, Subhransu MajiICCV 2021 · 18 citations
- Unsupervised Learning of Object Landmarks via Self-Training CorrespondenceDimitrios Mallis, Enrique Sanchez, Matthew Bell, Georgios TzimiropoulosNeurIPS 2020 · 20 citations
- Viewpoint Invariant Dense Matching for Visual GeolocalizationGabriele Moreno Berton, Carlo Masone, Valerio Paolicelli, Barbara CaputoICCV 2021 · 48 citations
- Learning Rotation-Equivariant Features for Visual CorrespondenceJongmin Lee, Byungjin Kim, Seungwook Kim, Minsu ChoCVPR 2023
