Alchemist: Parametric Control of Material Properties with Diffusion Models
Prafull Sharma, Varun Jampani, Yuanzhen Li, Xuhui Jia, Dmitry Lagun, Frédo Durand, Bill Freeman, Mark J. Matthews
Abstract
We propose a method to control material attributes of objects like roughness, metallic, albedo, and transparency in real images. Our method capitalizes on the generative prior of text-to-image models known for photorealism, employing a scalar value and instructions to alter low-level material properties. Addressing the lack of datasets with controlled material attributes, we generated an object-centric synthetic dataset with physically-based materials. Finetuning a modified pre-trained text-to-image model on this *This research was performed while Prafull Sharma was at Google. † Varun Jampani is now at Stability AI. synthetic dataset enables us to edit material properties in real-world images while preserving all other attributes. We show the potential application of our model to material edited NeRFs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers20
- RGB↔X: Image decomposition and synthesis using material- and lighting-aware diffusion modelsZheng Zeng, Valentin Deschaintre, Iliyan Georgiev, Yannick Hold-Geoffroy et al.SIGGRAPH 2024 · 61 citations
- DiLightNet: Fine-grained Lighting Control for Diffusion-based Image GenerationChong Zeng, Yue Dong, Pieter Peers, Youkang Kong et al.SIGGRAPH 2024 · 41 citations
- TextureDreamer: Image-Guided Texture Synthesis through Geometry-Aware DiffusionYu-Ying Yeh, Jia-Bin Huang, Changil Kim, Lei Xiao et al.CVPR 2024 · 31 citations
- IntrinsiX: High-Quality PBR Generation using Image PriorsPeter Kocsis, Lukas Höllein, Matthias NießnerNeurIPS 2025 · 20 citations
- Kontinuous Kontext: Continuous Strength Control for Instruction-based Image EditingRishubh Parihar, Or Patashnik, Daniil Ostashev, Venkatesh Babu Radhakrishnan et al.CVPR 2026 · 15 citations
Builds on43
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam et al.ICML 2022 · 4,691 citations
Related papers
- MARBLE: Material Recomposition and Blending in CLIP-SpaceTa Ying Cheng, Prafull Sharma, Mark Boss, Varun JampaniCVPR 2025
- PhyS-EdiT: Physics-aware Semantic Image Editing with Text DescriptionZiqi Cai, Shuchen Weng, Yifei Xia, Boxin ShiCVPR 2025
- Alterbute: Editing Intrinsic Attributes of Objects in ImagesTal Reiss, Daniel Winter, Matan Cohen, Alex Rav-Acha et al.ICML 2026
- MaterialMVP: Illumination-Invariant Material Generation via Multi-View PBR DiffusionZebin He, Mingxin Yang, Shuhui Yang, Yixuan Tang et al.ICCV 2025 · 3 citations
- PhotoMat: A Material Generator Learned from Single Flash PhotosXilong Zhou, Milos Hasan, Valentin Deschaintre, Paul Guerrero et al.SIGGRAPH 2023 · 31 citations
