Locality-Aware Generalizable Implicit Neural Representation
Doyup Lee, Chiheon Kim, Minsu Cho, Wook-Shin Han
Abstract
Generalizable implicit neural representation (INR) enables a single continuous function, i.e., a coordinate-based neural network, to represent multiple data instances by modulating its weights or intermediate features using latent codes. However, the expressive power of the state-of-the-art modulation is limited due to its inability to localize and capture fine-grained details of data entities such as specific pixels and rays. To address this issue, we propose a novel framework for generalizable INR that combines a transformer encoder with a locality-aware INR decoder. The transformer encoder predicts a set of latent tokens from a data instance to encode local information into each latent token. The locality-aware INR decoder extracts a modulation vector by selectively aggregating the latent tokens via cross-attention for a coordinate input and then predicts the output by progressively decoding with coarse-to-fine modulation through multiple frequency bandwidths. The selective token aggregation and the multi-band feature modulation enable us to learn locality-aware representation in spatial and spectral aspects, respectively. Our framework significantly outperforms previous generalizable INRs and validates the usefulness of the locality-aware latents for downstream tasks such as image generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6f81742b-58aa-48fe-98d1-504c8c293cf6Cited by top-tier papers12
- AROMA: Preserving Spatial Structure for Latent PDE Modeling with Local Neural FieldsLouis Serrano, Thomas X. Wang, Etienne Le Naour, Jean-Noël Vittaut et al.NeurIPS 2024 · 47 citations
- Generalizable Implicit Motion Modeling for Video Frame InterpolationZujin Guo, Wei Li, Chen Change LoyNeurIPS 2024 · 24 citations
- Pre-training Sequence, Structure, and Surface Features for Comprehensive Protein Representation LearningYouhan Lee, Hasun Yu, Jaemyung Lee, Jaehoon KimICLR 2024 · 22 citations
- CryoHype: Reconstructing a thousand cryo-EM structures with transformer-based hypernetworksJeffrey Gu, Minkyu Jeon, Ambri Ma, Serena Yeung-Levy et al.CVPR 2026 · 1 citation
- DVI: A Derivative-based Vision Network for INRRunzhao Yang, Xiaolong Wu, Zhihong Zhang, Fabian Zhang et al.ICML 2025
Builds on21
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
Related papers
- Generalizable Implicit Neural Representations via Instance Pattern ComposersChiheon Kim, Doyup Lee, Saehoon Kim, Minsu Cho et al.CVPR 2023
- Adversarial Generation of Continuous ImagesIvan Skorokhodov, Savva Ignatyev, Mohamed ElhoseinyCVPR 2021
- Cascaded Local Implicit Transformer for Arbitrary-Scale Super-ResolutionHao-Wei Chen, Yu-Syuan Xu, Min-Fong Hong, Yi-Min Tsai et al.CVPR 2023
- Learning Spatially Collaged Fourier Bases for Implicit Neural RepresentationJason Chun Lok Li, Chang Liu, Binxiao Huang, Ngai WongAAAI 2024 · 14 citations
- Implicit Neural Representations with Levels-of-ExpertsZekun Hao, Arun Mallya, Serge J. Belongie, Ming-Yu LiuNeurIPS 2022 · 27 citations
