G-IR: Geometric Image Representation for Learning
Xin Chen, Qi Zhao, Wei Zeng, Zongben Xu
摘要
Images are generally represented by pixel intensities or color values, which are usually used as direct inputs for learning. This study innovatively proposes a geometric image representation method and refreshes the general learning model (e.g., autoencoder) in the diffeomorphic space. Based on the theory of geometric optimal transport and quasiconformal mapping, we equivalently transform the intensity representation into a shape representation. The image space becomes a diffeomorphic space, where any image can be uniquely represented as a Beltrami coefficient function defined on a uniform grid reference, and vice versa. This innovative geometric image representation (G-IR) captures the fine-grained structure inherent in the entire image, which is different from the traditional feature extraction that focuses on the internal geometric objects of the image (such as boundaries and axes). The diffeomorphic property preserves structure in the generation process, which is very necessary in the field of real physics. It can be assembled into existing pipelines as a plug-in, providing structure-preserving properties for the entire framework. Experiments on image restoration and interpolation validated the high efficiency, efficacy and applicability of the G-IR method, demonstrating its superior performance compared to common pixel-level image appearance representations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- From data to functa: Your data point is a function and you can treat it like oneEmilien Dupont, Hyunjik Kim, S. M. Ali Eslami, Danilo Jimenez Rezende 等ICML 2022 · 被引用 209 次
- Diverse Plausible 360-Degree Image Outpainting for Efficient 3DCG Background CreationNaofumi Akimoto, Yuhi Matsuo, Yoshimitsu AokiCVPR 2022 · 被引用 34 次
- Large Images Are Gaussians: High-Quality Large Image Representation with Levels of 2D Gaussian SplattingLingting Zhu, Guying Lin, Jinnan Chen, Xinjie Zhang 等AAAI 2025 · 被引用 23 次
相关 Paper
- Geo-SIC: Learning Deformable Geometric Shapes in Deep Image ClassifiersJian Wang, Miaomiao ZhangNeurIPS 2022 · 被引用 12 次
- Regularized Autoencoders for Isometric Representation LearningYonghyeon Lee, Sangwoong Yoon, Minjun Son, Frank Chongwoo ParkICLR 2022 · 被引用 46 次
- Unified Latent Space for Understanding and Generation via Semantic Auto-encoderXiaojie Li, Yang Zhao, Ming Li, Yancheng Zhang 等CVPR 2026
- Revisiting Self-Similarity: Structural Embedding for Image RetrievalSeongwon Lee, Suhyeon Lee, Hongje Seong, Euntai KimCVPR 2023
- Geometry DistributionsBiao Zhang, Jing Ren, Peter WonkaICCV 2025 · 被引用 2 次
