Layout-Guided Novel View Synthesis From a Single Indoor Panorama
Jiale Xu, Jia Zheng, Yanyu Xu, Rui Tang, Shenghua Gao
Abstract
Existing view synthesis methods mainly focus on the perspective images and have shown promising results. However, due to the limited field-of-view of the pinhole camera, the performance quickly degrades when large camera movements are adopted. In this paper, we make the first attempt to generate novel views from a single indoor panorama and take the large camera translations into consideration. To tackle this challenging problem, we first use Convolutional Neural Networks (CNNs) to extract the deep features and estimate the depth map from the source-view image. Then, we leverage the room layout prior, a strong structural constraint of the indoor scene, to guide the generation of target views. More concretely, we estimate the room layout in the source view and transform it into the target viewpoint as guidance. Meanwhile, we also constrain the room layout of the generated target-view images to enforce geometric consistency. To validate the effectiveness of our method, we further build a large-scale photorealistic dataset containing both small and large camera translations. The experimental results on our challenging dataset demonstrate that our method achieves stateof-the-art performance. The project page is at https: //github.com/bluestyle97/PNVS .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8edbdac0-f157-44b0-9896-e2e86a0758d8Cited by top-tier papers6
- Text2Room: Extracting Textured 3D Meshes from 2D Text-to-Image ModelsLukas Höllein, Ang Cao, Andrew Owens, Justin Johnson et al.ICCV 2023 · 292 citations
- Taming Stable Diffusion for Text to 360° Panorama Image GenerationCheng Zhang, Qianyi Wu, Camilo Cruz Gambardella, Xiaoshui Huang et al.CVPR 2024 · 27 citations
- HORIZON: High-Resolution Semantically Controlled Panorama SynthesisKun Yan, Lei Ji, Chenfei Wu, Jian Liang et al.AAAI 2024 · 3 citations
- SO(3)-Equivariant ViT-Adapter for Data-Efficient Zero-Shot Sim-to-Real Indoor Panoramic Depth EstimationZiyan He, Qiudan Zhang, Lin Ma, Xu WangCVPR 2026
- Look Beyond: Two-Stage Scene View Generation via Panorama and Video DiffusionXueyang Kang, Zhengkang Xiang, Zezheng Zhang, Kourosh KhoshelhamACM MM 2025
Builds on13
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen et al.ICCV 2019 · 1,990 citations
- Coherent Semantic Attention for Image InpaintingHongyu Liu, Bin Jiang, Yi Xiao, Chao YangICCV 2019 · 395 citations
- StructureFlow: Image Inpainting via Structure-Aware Appearance FlowYurui Ren, Xiaoming Yu, Ruonan Zhang, Thomas H. Li et al.ICCV 2019 · 356 citations
- Extreme View SynthesisInchang Choi, Orazio Gallo, Alejandro J. Troccoli, Min H. Kim et al.ICCV 2019 · 207 citations
- Progressive Reconstruction of Visual Structure for Image InpaintingJingyuan Li, Fengxiang He, Lefei Zhang, Bo Du et al.ICCV 2019 · 151 citations
Related papers
- Look Outside the Room: Synthesizing A Consistent Long-Term 3D Scene Video from A Single ImageXuanchi Ren, Xiaolong WangCVPR 2022 · 42 citations
- Learning Object Context for Novel-view Scene Layout GenerationXiaotian Qiao, Gerhard P. Hancke, Rynson W. H. LauCVPR 2022 · 6 citations
- Generative View Synthesis: From Single-view Semantics to Novel-view ImagesTewodros Amberbir Habtegebrial, Varun Jampani, Orazio Gallo, Didier StrickerNeurIPS 2020 · 20 citations
- P2I-NET: Mapping Camera Pose to Image via Adversarial Learning for New View Synthesis in Real Indoor EnvironmentsXujie Kang, Kanglin Liu, Jiang Duan, Yuanhao Gong et al.ACM MM 2023 · 2 citations
- MultiDiff: Consistent Novel View Synthesis from a Single ImageNorman Müller, Katja Schwarz, Barbara Rössle, Lorenzo Porzi et al.CVPR 2024 · 14 citations
