Neural Matching Fields: Implicit Representation of Matching Fields for Visual Correspondence
Sunghwan Hong, Jisu Nam, Seokju Cho, Susung Hong, Sangryul Jeon, Dongbo Min, Seungryong Kim
Abstract
Existing pipelines of semantic correspondence commonly include extracting high-level semantic features for the invariance against intra-class variations and background clutters. This architecture, however, inevitably results in a low-resolution matching field that additionally requires an ad-hoc interpolation process as a post-processing for converting it into a high-resolution one, certainly limiting the overall performance of matching results. To overcome this, inspired by recent success of implicit neural representation, we present a novel method for semantic correspondence, called Neural Matching Field (NeMF). However, complicacy and high-dimensionality of a 4D matching field are the major hindrances, which we propose a cost embedding network to process a coarse cost volume to use as a guidance for establishing high-precision matching field through the following fully-connected network. Nevertheless, learning a high-dimensional matching field remains challenging mainly due to computational complexity, since a naive exhaustive inference would require querying from all pixels in the 4D space to infer pixel-wise correspondences. To overcome this, we propose adequate training and inference procedures, which in the training phase, we randomly sample matching candidates and in the inference phase, we iteratively performs PatchMatch-based inference and coordinate optimization at test time. With these combined, competitive results are attained on several standard benchmarks for semantic correspondence. Code and pre-trained weights are available at https://ku-cvlab.github.io/NeMF/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e83a92b1-a02a-45b8-87a1-1e64b0c45e4bCited by top-tier papers16
- Emergent Temporal Correspondences from Video Diffusion TransformersJisu Nam, Soowon Son, Dahyun Chung, Jiyoung Kim et al.NeurIPS 2025 · 30 citations
- DreamMatcher: Appearance Matching Self-Attention for Semantically-Consistent Text-to-Image PersonalizationJisu Nam, Heesu Kim, DongJae Lee, Siyoon Jin et al.CVPR 2024 · 21 citations
- Emergent Outlier View Rejection in Visual Geometry Grounded TransformersJisang Han, Sunghwan Hong, Jaewoo Jung, Wooseok Jang et al.CVPR 2026 · 19 citations
- Unifying Feature and Cost Aggregation with Transformers for Semantic and Visual CorrespondenceSunghwan Hong, Seokju Cho, Seungryong Kim, Stephen LinICLR 2024 · 16 citations
- Enhancing 3D Reconstruction for Dynamic ScenesJisang Han, Honggyu An, Jaewoo Jung, Takuya Narihira et al.NeurIPS 2025 · 11 citations
Builds on31
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- On the Variance of the Adaptive Learning Rate and BeyondLiyuan Liu, Haoming Jiang, Pengcheng He, Weizhu Chen et al.ICLR 2020 · 2,210 citations
Related papers
- ConvMatch: Rethinking Network Design for Two-View Correspondence LearningShihua Zhang, Jiayi MaAAAI 2023 · 57 citations
- Semantic-Aware Implicit Template Learning via Part Deformation ConsistencySihyeon Kim, Juyeon Ko, Minseok Joo, Juhan Cha et al.ICCV 2023 · 3 citations
- NeMF: Neural Motion Fields for Kinematic AnimationChengan He, Jun Saito, James Zachary, Holly E. Rushmeier et al.NeurIPS 2022 · 84 citations
- Pixel-Level Semantic Correspondence Through Layout-Aware Representation Learning and Multi-Scale Matching IntegrationYixuan Sun, Zhangyue Yin, Haibo Wang, Yan Wang et al.CVPR 2024 · 5 citations
- PatchMatch-Based Neighborhood Consensus for Semantic CorrespondenceJae Yong Lee, Joseph DeGol, Victor Fragoso, Sudipta N. SinhaCVPR 2021
