Scene-Level Appearance Transfer with Semantic Correspondences
Liyuan Zhu, Shengqu Cai, Shengyu Huang, Gordon Wetzstein, Naji Khosravan, Iro Armeni
摘要
We introduce ReStyle3D, a novel framework for scene-level appearance transfer from a single style image to a real-world scene represented by multiple views. The method combines explicit semantic correspondences with multi-view consistency to achieve precise and coherent stylization. Unlike conventional stylization methods that apply a reference style globally, ReStyle3D uses open-vocabulary segmentation to establish dense, instance-level correspondences between the style and real-world images. This ensures that each object is stylized with semantically matched textures. ReStyle3D first transfers the style to a single view using a training-free semantic-attention mechanism in a diffusion model. It then lifts the stylization to additional views via a learned warp-and-refine network guided by monocular depth and pixel-wise correspondences. Experiments show that ReStyle3D consistently outperforms prior methods in structure preservation, perceptual style similarity, and multi-view coherence. User studies further validate its ability to produce photo-realistic, semantically faithful results. Our code, pretrained models, and dataset will be publicly released, to support new applications in interior design, virtual staging, and 3D-consistent stylization. Project page and code at https://restyle3d.github.io/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- GuideFlow3D: Optimization-Guided Rectified Flow For Appearance TransferSayan Deb Sarkar, Sinisa Stekovic, Vincent Lepetit, Iro ArmeniNeurIPS 2025 · 被引用 3 次
- Shape-of-You: Fused Gromov-Wasserstein Optimal Transport for Semantic Correspondence in-the-WildJiin Im, Sisung Liu, Je Hyeong HongCVPR 2026
它引用的顶会 Paper37
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann 等ICLR 2024 · 被引用 4,569 次
相关 Paper
- DiffStyle3D: Consistent 3D Gaussian Stylization via Attention OptimizationYitong Yang, Yinglin Wang, Xuexin Liu, Jing Wang 等ICML 2026 · 被引用 2 次
- CoCoDiff: Correspondence-Consistent Diffusion Model for Fine-grained Style TransferWenbo Nie, Zixiang Li, Renshuai Tao, Bin WU 等ICLR 2026 · 被引用 2 次
- FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance FieldsGeonU Kim, Kim Youwang, Tae-Hyun OhAAAI 2024 · 被引用 12 次
- Stylos: Multi-View 3D Stylization with Single-Forward Gaussian SplattingHanzhou Liu, Jia Huang, Mi Lu, Srikanth Saripalli 等ICLR 2026 · 被引用 4 次
- PNeSM: Arbitrary 3D Scene Stylization via Prompt-Based Neural Style MappingJiafu Chen, Wei Xing, Jiakai Sun, Tianyi Chu 等AAAI 2024 · 被引用 2 次
