Long-Term Temporally Consistent Unpaired Video Translation from Simulated Surgical 3D Data
Dominik Rivoir, Micha Pfeiffer, Reuben Docea, Fiona R. Kolbinger, Carina Riediger, Jürgen Weitz, Stefanie Speidel
Abstract
Research in unpaired video translation has mainly focused on short-term temporal consistency by conditioning on neighboring frames. However for transfer from simulated to photorealistic sequences, available information on the underlying geometry offers potential for achieving global consistency across views. We propose a novel approach which combines unpaired image translation with neural rendering to transfer simulated to photorealistic surgical abdominal scenes. By introducing global learnable textures and a lighting-invariant view-consistency loss, our method produces consistent translations of arbitrary views and thus enables long-term consistent video synthesis. We design and test our model to generate video sequences from minimally-invasive surgical abdominal scenes. Because labeled data is often limited in this domain, photorealistic data where ground truth information from the simulated domain is preserved is especially relevant. By extending existing image-based methods to view-consistent videos, we aim to impact the applicability of simulated training and evaluation environments for surgical applications. Code and data: http://opencas.dkfz.de/video-sim2real.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8275d286-917a-4885-a5ab-6b4177600bc0Builds on4
- Few-Shot Unsupervised Image-to-Image TranslationMing-Yu Liu, Xun Huang, Arun Mallya, Tero Karras et al.ICCV 2019 · 668 citations
- Learning temporal coherence via self-supervision for GAN-based video generationMengyu Chu, You Xie, Jonas Mayer, Laura Leal-Taixé et al.SIGGRAPH 2020 · 198 citations
- Augmenting Colonoscopy Using Extended and Directional CycleGAN for Lossy Image TranslationShawn Mathew, Saad Nadeem, Sruti Kumari, Arie E. KaufmanCVPR 2020
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten et al.CVPR 2020
Related papers
- TexPose: Neural Texture Learning for Self-Supervised 6D Object Pose EstimationHanzhi Chen, Fabian Manhardt, Nassir Navab, Benjamin BusamCVPR 2023
- Putting NeRF on a Diet: Semantically Consistent Few-Shot View SynthesisAjay Jain, Matthew Tancik, Pieter AbbeelICCV 2021 · 615 citations
- Look Outside the Room: Synthesizing A Consistent Long-Term 3D Scene Video from A Single ImageXuanchi Ren, Xiaolong WangCVPR 2022 · 42 citations
- The Spatially-Correlative Loss for Various Image Translation TasksChuanxia Zheng, Tat-Jen Cham, Jianfei CaiCVPR 2021
- Geometry-Aware Satellite-to-Ground Image Synthesis for Urban AreasXiaohu Lu, Zuoyue Li, Zhaopeng Cui, Martin R. Oswald et al.CVPR 2020
