CleAR: Robust Context-Guided Generative Lighting Estimation for Mobile Augmented Reality
Yiqin Zhao, Mallesham Dasari, Tian Guo
摘要
High-quality environment lighting is essential for creating immersive mobile augmented reality (AR) experiences. However, achieving visually coherent estimation for mobile AR is challenging due to several key limitations in AR device sensing capabilities, including low camera FoV and limited pixel dynamic ranges. Recent advancements in generative AI, which can generate high-quality images from different types of prompts, including texts and images, present a potential solution for high-quality lighting estimation. Still, to effectively use generative image diffusion models, we must address two key limitations of content quality and slow inference. In this work, we design and implement a generative lighting estimation system called CleAR that can produce high-quality, diverse environment maps in the format of 360 • HDR images. Specifically, we design a two-step generation pipeline guided by AR environment context data to ensure the output aligns with the physical environment's visual context and color appearance. To improve the estimation robustness under different lighting conditions, we design a real-time refinement component to adjust lighting estimation results on AR devices. To train and test our generative models, we curate a large-scale environment lighting estimation dataset with diverse lighting conditions. Through a combination of quantitative and qualitative evaluations, we show that CleAR outperforms state-of-the-art lighting estimation methods on both estimation accuracy, latency, and robustness, and is rated by 31 participants as producing better renderings for most virtual objects. For example, CleAR achieves 51% to 56% accuracy improvement on virtual object renderings across objects of three distinctive types of materials and reflective properties. CleAR produces lighting estimates of comparable or better quality in just 3.2 seconds-over 110X faster than state-of-the-art methods. Moreover, CleAR supports real-time refinement of lighting estimation results, ensuring robust and timely updates for AR applications.
CCS Concepts: • Computing methodologies → Mixed / augmented reality; • Human-centered computing → Ubiquitous and mobile computing systems and tools.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper17
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- UniPC: A Unified Predictor-Corrector Framework for Fast Sampling of Diffusion ModelsWenliang Zhao, Lujia Bai, Yongming Rao, Jie Zhou 等NeurIPS 2023 · 被引用 537 次
- Deep Parametric Indoor Lighting EstimationMarc-André Gardner, Yannick Hold-Geoffroy, Kalyan Sunkavalli, Christian Gagné 等ICCV 2019 · 被引用 155 次
相关 Paper
- Rendering-Aware HDR Environment Map Prediction from a Single ImageJun-Peng Xu, Chenyu Zuo, Fang-Lue Zhang, Miao WangAAAI 2022 · 被引用 15 次
- HDR Environment Map Estimation for Real-Time Augmented RealityGowri Somanath, Daniel KurzCVPR 2021
- EverLight: Indoor-Outdoor Editable HDR Lighting EstimationMohammad Reza Karimi Dastjerdi, Jonathan Eisenmann, Yannick Hold-Geoffroy, Jean-François LalondeICCV 2023 · 被引用 41 次
- GSHeadRelight: Fast Relightability for 3D Gaussian Head SynthesisHenglei Lv, Bailin Deng, Jianzhu Guo, Xiaoqiang Liu 等SIGGRAPH 2025
- LuxDiT: Lighting Estimation with Video Diffusion TransformerRuofan Liang, Kai He, Zan Gojcic, Igor Gilitschenski 等NeurIPS 2025 · 被引用 20 次
