Relighting as a Probe of Visual Priors via Augmented Latent Intrinsics
Xiaoyan Xing, Xiao Zhang, Sezer Karaoglu, Theo Gevers, Anand Bhattad
摘要
Generative relighting with different visual representation features Figure 1 . Stronger Semantic Encoders Can Harm Relighting Performance. Left: Visual comparison on a scene with complex specular materials. The task is to relight the input image (top-left) using the target illumination (bottom-left), which requires moving specular highlights from left to right, as indicated by the chrome sphere. While features from semantic encoders (CLIP, DINO) fail to reproduce realistic highlights, the MAE plausibly moves the highlight but blurs fine details, such as text labels. Our method (top-right), which combines features from RADIO (a pretrained model; distilled from many vision encoders) with latent intrinsics, closely matches the ground truth. Right: Quantitative analysis reveals a trade-off: for most encoders optimized for pure semantics, relighting quality (PSNR) is inversely correlated with recognition performance (ImageNet-1K linear probing as reported in the original papers.). Our approach breaks this trend, achieving high performance on both tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Diffusion Transformers with Representation AutoencodersBoyang Zheng, Nanye Ma, Shengbang Tong, Saining XieICLR 2026 · 被引用 288 次
相关 Paper
- Latent Intrinsics Emerge from Training to RelightXiao Zhang, William Gao, Seemandhar Jain, Michael Maire 等NeurIPS 2024 · 被引用 21 次
- LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene RelightingXiaoyan Xing, Konrad Groh, Sezer Karaoglu, Theo Gevers 等CVPR 2025
- Physically Controllable Relighting of PhotographsChris Careaga, Yagiz AksoySIGGRAPH 2025 · 被引用 3 次
- TokenLight: Precise Lighting Control in Images using Attribute TokensSumit Chaturvedi, Yannick Hold-Geoffroy, Mengwei Ren, Jingyuan Liu 等CVPR 2026 · 被引用 1 次
- IntrinsicDiffusion: Joint Intrinsic Layers from Latent Diffusion ModelsJundan Luo, Duygu Ceylan, Jae Shin Yoon, Nanxuan Zhao 等SIGGRAPH 2024 · 被引用 18 次
