M3ashy: Multi-Modal Material Synthesis via Hyperdiffusion
Chenliang Zhou, Zheyuan Hu, Alejandro Sztrajman, Yancheng Cai, Yaru Liu, Cengiz Öztireli
Abstract
High-quality material synthesis is essential for replicating complex surface properties to create realistic scenes. Despite advances in the generation of material appearance based on analytic models, the synthesis of real-world measured BRDFs remains largely unexplored. To address this challenge, we propose M 3 ashy, a novel multi-modal material synthesis framework based on hyperdiffusion. M 3 ashy enables highquality reconstruction of complex real-world materials by leveraging neural fields as a compact continuous representation of BRDFs. Furthermore, our multi-modal conditional hyperdiffusion model allows for flexible material synthesis conditioned on material type, natural language descriptions, or reference images, providing greater user control over material generation. To support future research, we contribute two new material datasets and introduce two BRDF distributional metrics for more rigorous evaluation. We demonstrate the effectiveness of M 3 ashy through extensive experiments, including a novel statistics-based constrained synthesis, which enables the generation of materials of desired categories.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 387e5fa1-581b-4ce9-8e53-4e6e32624ee2Cited by top-tier papers2
- Generative Modeling of Weights: Generalization or Memorization?Boya Zeng, Yida Yin, Zhiqiu Xu, Zhuang LiuCVPR 2026 · 12 citations
- Toward Richer Material Generation via Procedural Data EnhancementYunchen Yu, Jacob Munkberg, Jon Hasselgren, Chris Cummings et al.SIGGRAPH 2026
Builds on11
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- PointFlow: 3D Point Cloud Generation With Continuous Normalizing FlowsGuandao Yang, Xun Huang, Zekun Hao, Ming-Yu Liu et al.ICCV 2019 · 794 citations
Related papers
- PhotoMat: A Material Generator Learned from Single Flash PhotosXilong Zhou, Milos Hasan, Valentin Deschaintre, Paul Guerrero et al.SIGGRAPH 2023 · 31 citations
- TensoSDF: Roughness-aware Tensorial Representation for Robust Geometry and Material ReconstructionJia Li, Lu Wang, Lei Zhang, Beibei WangSIGGRAPH 2024 · 22 citations
- SViM3D: Stable Video Material Diffusion for Single Image 3D GenerationAndreas Engelhardt, Mark Boss, Vikram Voleti, Chun-Han Yao et al.ICCV 2025 · 1 citation
- Learning Neural Exposure Fields for View SynthesisMichael Niemeyer, Fabian Manhardt, Marie-Julie Rakotosaona, Michael Oechsle et al.NeurIPS 2025 · 6 citations
- Neural Layered BRDFsJiahui Fan, Beibei Wang, Milos Hasan, Jian Yang et al.SIGGRAPH 2022 · 31 citations
