DiffStyle3D: Consistent 3D Gaussian Stylization via Attention Optimization
Yitong Yang, Yinglin Wang, Xuexin Liu, Jing Wang, Hao Dou, Changshuo Wang, Shuting He
摘要
3D style transfer enables the creation of visually expressive 3D content, enriching the visual appearance of 3D scenes and objects. However, existing VGG- and CLIP-based methods struggle to model multi-view consistency within the model itself, while diffusion-based approaches can capture such consistency but rely on denoising directions, leading to unstable training. To address these limitations, we propose DiffStyle3D, a novel diffusion-based paradigm for 3DGS style transfer that directly optimizes in the latent space. Specifically, we introduce an Attention-Aware Loss that performs style transfer by aligning style features in the self-attention space, while preserving original content through content feature alignment. Inspired by the geometric invariance of 3D stylization, we propose a Geometry-Guided Multi-View Consistency method that integrates geometric information into self-attention to enable cross-view correspondence modeling. Based on geometric information, we additionally construct a geometry-aware mask to prevent redundant optimization in overlapping regions across views, which further improves multi-view consistency. Extensive experiments show that DiffStyle3D outperforms state-of-the-art methods, achieving higher stylization quality and visual realism. The code is available at https://github.com/yangyt46/DiffStyle3D.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper23
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 被引用 5,234 次
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann 等ICLR 2024 · 被引用 4,569 次
相关 Paper
- FantasyStyle: Controllable Stylized Distillation for 3D Gaussian SplattingYitong Yang, Yinglin Wang, Changshuo Wang, Huajie Wang 等AAAI 2026 · 被引用 2 次
- Scene-Level Appearance Transfer with Semantic CorrespondencesLiyuan Zhu, Shengqu Cai, Shengyu Huang, Gordon Wetzstein 等SIGGRAPH 2025 · 被引用 3 次
- Stylos: Multi-View 3D Stylization with Single-Forward Gaussian SplattingHanzhou Liu, Jia Huang, Mi Lu, Srikanth Saripalli 等ICLR 2026 · 被引用 4 次
- Tune-Your-Style: Intensity-Tunable 3D Style Transfer with Gaussian SplattingYian Zhao, Rushi Ye, Ruochong Zheng, Zesen Cheng 等ICCV 2025 · 被引用 3 次
- GeoVideo: Introducing Geometric Regularization into Video Generation ModelYunpeng Bai, Shaoheng Fang, Chaohui Yu, Fan Wang 等NeurIPS 2025 · 被引用 18 次
