Light of Normals: Unified Feature Representation for Universal Photometric Stereo
Houyuan Chen, Hong Li, Chongjie Ye, Zhaoxi Chen, Bohan Li, Shaocong Xu, Xianda Guo, Xuhui Liu, Yikai Wang, Baochang Zhang, Satoshi Ikehata, Boxin Shi
Abstract
Universal photometric stereo (PS) is defined by two factors: it must (i) operate under arbitrary, unknown lighting conditions and (ii) avoid reliance on specific illumination models. Despite progress (e.g., SDM UniPS), two challenges remain. First, current encoders cannot guarantee that illumination and normal information are decoupled. To enforce decoupling, we introduce LINO UniPS with two key components: (i) Light Register Tokens with light alignment supervision to aggregate point, direction, and environment lights; (ii) Interleaved Attention Block featuring global cross-image attention that takes all lighting conditions together so the encoder can factor out lighting while retaining normal-related evidence. Second, high-frequency geometric details are easily lost. We address this with (i) a Wavelet-based Dual-branch Architecture and (ii) a Normal-gradient Perception Loss. These techniques yield a unified feature space in which lighting is explicitly represented by register tokens, while normal details are preserved via wavelet branch. We further introduce PS-Verse, a large-scale synthetic dataset graded by geometric complexity and lighting diversity, and adopt curriculum training from simple to complex scenes. Extensive experiments show new state-of-the-art results on public benchmarks (e.g., DiLiGenT, Luces), stronger generalization to real materials, and improved efficiency; ablations confirm that Light Register Tokens + Interleaved Attention Block drive better feature decoupling, while Wavelet-based Dual-branch Architecture + Normal-gradient Perception Loss recover finer details.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 512350e3-0bf9-4afb-9792-2aa784119b37Cited by top-tier papers4
- PAGE-4D: Disentangled Pose and Geometry Estimation for VGGT-4D PerceptionKaichen Zhou, Yuhan Wang, Grace Chen, Gaspard Beaudouin et al.ICLR 2026 · 12 citations
- NeAR: Coupled Neural Asset-Renderer StackHong Li, Chongjie Ye, Houyuan Chen, Weiqing Xiao et al.CVPR 2026 · 4 citations
- Geometry Meets Light: Leveraging Geometric Priors for Universal Photometric Stereo Under Limited Multi-Illumination CuesKing-Man Tam, Satoshi Ikehata, Yuta Asano, Zhaoyi An et al.AAAI 2026
- Relit-LiVE: Relight Video by Jointly Learning Environment VideoWeiqing Xiao, Hong Li, Xiuyu Yang, Houyuan Chen et al.SIGGRAPH 2026
Builds on17
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Point Transformer V2: Grouped Vector Attention and Partition-based PoolingXiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu et al.NeurIPS 2022 · 924 citations
- Vision Transformers Need RegistersTimothée Darcet, Maxime Oquab, Julien Mairal, Piotr BojanowskiICLR 2024 · 769 citations
- Towards Large-Scale 3D Representation Learning with Multi-Dataset Point Prompt TrainingXiaoyang Wu, Zhuotao Tian, Xin Wen, Bohao Peng et al.CVPR 2024 · 39 citations
Related papers
- Universal Photometric Stereo Network using Global Lighting ContextsSatoshi IkehataCVPR 2022 · 22 citations
- Spin-UP: Spin Light for Natural Light Uncalibrated Photometric StereoZongrui Li, Zhan Lu, Haojie Yan, Boxin Shi et al.CVPR 2024
- Learning Efficient Photometric Feature Transform for Multi-view StereoKaizhang Kang, Cihui Xie, Ruisheng Zhu, Xiaohe Ma et al.ICCV 2021 · 3 citations
- DiLiGenT-Π: Photometric Stereo for Planar Surfaces with Rich Details - Benchmark Dataset and BeyondFeishi Wang, Jieji Ren, Heng Guo, Mingjun Ren et al.ICCV 2023 · 18 citations
- PX-NET: Simple and Efficient Pixel-Wise Training of Photometric Stereo NetworksFotios Logothetis, Ignas Budvytis, Roberto Mecca, Roberto CipollaICCV 2021 · 60 citations
