Parametric Shadow Control for Portrait Generation in Text-to-Image Diffusion Models
Haoming Cai, Tsung-Wei Huang, Shiv Gehlot, Brandon Y. Feng, Sachin Shah, Guan-Ming Su, Christopher A. Metzler
Abstract
Text-to-image diffusion models excel at generating diverse portraits, but lack intuitive shadow control. Existing editing approaches, as post-processing, struggle to offer effective manipulation across diverse styles. Additionally, these methods either rely on expensive real-world light-stage data collection or require extensive computational resources for training. To address these limitations, we introduce Shadow Director, a method that extracts and manipulates hidden shadow attributes within well-trained diffusion models. Our approach uses a small estimation network that requires only a few thousand synthetic images and hours of training-no costly real-world light-stage data needed. Shadow Director enables parametric and intuitive control over shadow shape, placement, and intensity during portrait generation while preserving artistic integrity and identity across diverse styles. Despite training only on synthetic data built on real-world identities, it generalizes effectively to generated portraits with diverse styles, making it a more accessible and resource-friendly solution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ce3d3453-ab9d-4384-9939-9c7a2db99e35Cited by top-tier papers3
- PhotoFramer: Multi-modal Image Composition InstructionZhiyuan You, Ke Wang, He Zhang, Xin Cai et al.CVPR 2026 · 8 citations
- DA-VAE: Plug-in Latent Compression for Diffusion via Detail AlignmentXin Cai, Zhiyuan You, Zhoutong Zhang, Tianfan XueCVPR 2026 · 3 citations
- Linear Image Generation by Synthesizing Exposure BracketsYuekun Dai, Zhoutong Zhang, Shangchen Zhou, Nanxuan ZhaoCVPR 2026
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Exploring CLIP for Assessing the Look and Feel of ImagesJianyi Wang, Kelvin C. K. Chan, Chen Change LoyAAAI 2023 · 1,208 citations
- DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic DataStephanie Fu, Netanel Tamir, Shobhita Sundaram, Lucy Chai et al.NeurIPS 2023 · 413 citations
Related papers
- Structure-Guided Diffusion Models for High-Fidelity Portrait Shadow RemovalWanchang Yu, Qing Zhang, Rongjia Zheng, Wei-Shi ZhengICCV 2025 · 1 citation
- SynthLight: Portrait Relighting with Diffusion Model by Learning to Re-render Synthetic FacesSumit Chaturvedi, Mengwei Ren, Yannick Hold-Geoffroy, Jingyuan Liu et al.CVPR 2025
- Foreground Harmonization and Shadow Generation for Composite ImageJing Zhou, Ziqi Yu, Zhongyun Bao, Gang Fu et al.ACM MM 2024 · 7 citations
- Detail-Preserving Latent Diffusion for Stable Shadow RemovalJiamin Xu, Yuxin Zheng, Zelong Li, Chi Wang et al.CVPR 2025
- PhotoApp: photorealistic appearance editing of head portraitsMallikarjun B. R., Ayush Tewari, Abdallah Dib, Tim Weyrich et al.SIGGRAPH 2021 · 11 citations
