Continuous Transformation Superposition for Visual Comfort Enhancement of Casual Stereoscopic Photography
Yuzhong Chen, Qijin Shen, Yuzhen Niu, Wenxi Liu
摘要
Casual stereoscopic photography allows ordinary users to create a stereoscopic photo using two photos taken casually by a monocular camera. The visual comfort of a casual stereoscopic photo can greatly affect its visual experience. In this paper, we present a novel visual comfort enhancement method for casual stereoscopic photography via reinforcement learning based on continuous transformation superposition. We consider the transformation, in a continuous transformation space, to transform each view as superpositions of several basic continuous transformations, enabling more subtle and flexible image transformation operations to approach better solutions. To achieve the continuous transformation superposition, we prepare a collection of continuous transformation models for translation, rotation, and perspective transformations. Then we train a policy model to determine an optimal transformation chain to recurrently handle both the geometric constraints and disparity adjustment, and thereby enhance the visual comfort of casual stereoscopic images. We further propose an attention-based stereo feature fusion module that enhances and integrates the binocular information between the left and right views. Experimental results on three datasets demonstrate that our proposed method achieves superior performance to state-of-the-art methods.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Recurrent Enhancement of Visual Comfort for Casual Stereoscopic PhotographyYuzhen Niu, Qingyang Zheng, Wenxi Liu, Wenzhong GuoIEEE VR 2020 · 被引用 2 次
- Visual Comfort Aware-Reinforcement Learning for Depth Adjustment of Stereoscopic 3D ImagesHak Gu Kim, Minho Park, Sangmin Lee, Seongyeop Kim 等AAAI 2021 · 被引用 11 次
- PatchMatch-RL: Deep MVS with Pixelwise Depth, Normal, and VisibilityJae Yong Lee, Joseph DeGol, Chuhang Zou, Derek HoiemICCV 2021 · 被引用 35 次
- GAIT: Generating Aesthetic Indoor Tours with Deep Reinforcement LearningDesai Xie, Ping Hu, Xin Sun, Sören Pirk 等ICCV 2023 · 被引用 9 次
- Recovering Physically Plausible Human-Object Interactions from Monocular VideosDingbang Huang, Etienne Vouga, Qixing Huang, Georgios PavlakosCVPR 2026
