SivsFormer: Parallax-Aware Transformers for Single-image-based View Synthesis
Chunlan Zhang, Chunyu Lin, Kang Liao, Lang Nie, Yao Zhao
摘要
Single-image-based view synthesis is significant for generating a 3D scene and gains increasing attention in recent years. However, this task is challenging as it requires inferring contents beyond what is immediately visible. Previous methods directly predict the unknown views using the convolutional neural networks, but the generated views suffer from visually unpleasant holes, deformations, and artifacts. In this paper, we propose a Single-image-based view synthesis transformer (named SivsFormer) for high-quality and realistic view synthesis. In particular, a warping and occlusion handing module is designed to reduce the influence of parallax on the network. Subsequently, a disparity alignment module captures the long-range information over the scene and ensures that pixels move in a geometrically correct manner with soft probabilistic disparity maps. Moreover, we present a parallax-aware loss function to improve the quality of the synthetic images, which explicitly quantifies the magnitude of parallaxes. We conduct extensive experiments on popular KITTI and Cityscapes datasets. Benefitting from the proposed parallax-aware transformer, our approach achieves superior performance in both quantitative and qualitative evaluations.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- CVSformer: Cross-View Synthesis Transformer for Semantic Scene CompletionHaotian Dong, Enhui Ma, Lubo Wang, Miaohui Wang 等ICCV 2023 · 被引用 12 次
- Geometry-Free View Synthesis: Transformers and no 3D PriorsRobin Rombach, Patrick Esser, Björn OmmerICCV 2021 · 被引用 115 次
- VIAFormer: Voxel-Image Alignment Transformer for High-Fidelity Voxel RefinementTiancheng Fang, Bowen Pan, Lingxi Chen, Jiangjing Lyu 等CVPR 2026 · 被引用 1 次
- VoxFormer: Sparse Voxel Transformer for Camera-Based 3D Semantic Scene CompletionYiming Li, Zhiding Yu, Christopher B. Choy, Chaowei Xiao 等CVPR 2023
- Dual-S3D: Hierarchical Dual-Path Selective SSM-CNN for High-Fidelity Implicit ReconstructionLuoxi Zhang, Pragyan Shrestha, Yu Zhou, Chun Xie 等ICCV 2025
