Relit-LiVE: Relight Video by Jointly Learning Environment Video
Weiqing Xiao, Hong Li, Xiuyu Yang, Houyuan Chen, Wenyi Li, Tianqi Liu, Shaocong Xu, Chongjie Ye, Hao Zhao, Beibei Wang
Abstract
Recent advances have shown that large-scale video diffusion models can be repurposed as neural renderers by first decomposing videos into intrinsic scene representations and then performing forward rendering under novel illumination. While promising, this paradigm fundamentally relies on accurate intrinsic decomposition, which remains highly unreliable for real-world videos and often leads to distorted appearances, broken materials, and accumulated temporal artifacts during relighting. In this work, we present Relit-LiVE , a novel video relighting framework that produces physically consistent, temporally stable results without requiring prior knowledge of camera pose. Our key insight is to explicitly introduce raw reference images into the rendering process, enabling the model to recover critical scene cues that are inevitably lost or corrupted in intrinsic representations. Furthermore, we propose a novel environment video prediction formulation that simultaneously generates relit videos and per-frame environment maps aligned with each camera viewpoint in a single diffusion process. This joint prediction enforces strong geometric–illumination alignment and naturally supports dynamic lighting and camera motion, significantly improving physical consistency in video relighting while easing the requirement of known per-frame camera pose. To further enhance generalization, we introduce two complementary training strategies: (i) latent-space interpolation between relighting and rendering outputs to synthesize diverse, photorealistic multi-illumination data, and (ii) a cycle-consistent self-supervised illumination learning scheme that enforces temporal lighting coherence without additional annotations. Extensive experiments demonstrate that Relit-LiVE consistently outperforms state-of-the-art video relighting and neural rendering methods across synthetic and real-world benchmarks. Beyond relighting, our framework naturally supports a wide range of downstream applications, including scene-level rendering, material editing, object insertion, and streaming video relighting. The Project is available at https://github.com/zhuxing0/Relit-LiVE.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cfe25800-b8e6-43e1-a43f-c1e6e546224aBuilds on22
- Neural Gaffer: Relighting Any Object via DiffusionHaian Jin, Yuan Li, Fujun Luan, Yuanbo Xiangli et al.NeurIPS 2024 · 112 citations
- SpatialVID: A Large-Scale Video Dataset with Spatial AnnotationsJiahao Wang, Yufeng Yuan, Rujie Zheng, Youtian Lin et al.CVPR 2026 · 72 citations
- RGB↔X: Image decomposition and synthesis using material- and lighting-aware diffusion modelsZheng Zeng, Valentin Deschaintre, Iliyan Georgiev, Yannick Hold-Geoffroy et al.SIGGRAPH 2024 · 61 citations
- NormalCrafter: Learning Temporally Consistent Normals from Video Diffusion PriorsYanrui Bin, Wenbo Hu, Haoyuan Wang, Xinya Chen et al.ICCV 2025 · 27 citations
- MatSynth: A Modern PBR Materials DatasetGiuseppe Vecchio, Valentin DeschaintreCVPR 2024 · 24 citations
Related papers
- UniRelight: Learning Joint Decomposition and Synthesis for Video RelightingKai He, Ruofan Liang, Jacob Munkberg, Jon Hasselgren et al.NeurIPS 2025 · 42 citations
- GR3EN: Generative Relighting for 3D EnvironmentsXiaoyan Xing, Philipp Henzler, Junhwa Hur, Runze Li et al.SIGGRAPH 2026
- Physically Controllable Relighting of PhotographsChris Careaga, Yagiz AksoySIGGRAPH 2025 · 3 citations
- Comprehensive Relighting: Generalizable and Consistent Monocular Human Relighting and HarmonizationJunying Wang, Jingyuan Liu, Xin Sun, Krishna Kumar Singh et al.CVPR 2025
- SViM3D: Stable Video Material Diffusion for Single Image 3D GenerationAndreas Engelhardt, Mark Boss, Vikram Voleti, Chun-Han Yao et al.ICCV 2025 · 1 citation
