Generative Adversarial Networks for Video-to-Video Domain Adaptation
Jiawei Chen, Yuexiang Li, Kai Ma, Yefeng Zheng
Abstract
Endoscopic videos from multicentres often have different imaging conditions, e.g., color and illumination, which make the models trained on one domain usually fail to generalize well to another. Domain adaptation is one of the potential solutions to address the problem. However, few of existing works focused on the translation of video-based data. In this work, we propose a novel generative adversarial network (GAN), namely VideoGAN, to transfer the video-based data across different domains. As the frames of a video may have similar content and imaging conditions, the proposed VideoGAN has an X-shape generator to preserve the intra-video consistency during translation. Furthermore, a loss function, namely color histogram loss, is proposed to tune the color distribution of each translated frame. Two colonoscopic datasets from different centres, i.e., CVC-Clinic and ETIS-Larib, are adopted to evaluate the performance of domain adaptation of our VideoGAN. Experimental results demonstrate that the adapted colonoscopic video generated by our VideoGAN can significantly boost the segmentation accuracy, i.e., an improvement of 5%, of colorectal polyps on multicentre datasets. As our VideoGAN is a general network architecture, we also evaluate its performance with the CamVid driving video dataset on the cloudy-to-sunny translation task. Comprehensive experiments show that the domain gap could be substantially narrowed down by our VideoGAN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on1
Related papers
- Augmenting Colonoscopy Using Extended and Directional CycleGAN for Lossy Image TranslationShawn Mathew, Saad Nadeem, Sruti Kumari, Arie E. KaufmanCVPR 2020
- Polymorphic-GAN: Generating Aligned Samples across Multiple Domains with Learned Morph MapsSeung Wook Kim, Karsten Kreis, Daiqing Li, Antonio Torralba et al.CVPR 2022 · 6 citations
- Unsupervised Image-to-Image Translation with Generative PriorShuai Yang, Liming Jiang, Ziwei Liu, Chen Change LoyCVPR 2022 · 51 citations
- Collaborative and Adversarial Learning of Focused and Dispersive Representations for Semi-supervised Polyp SegmentationHuisi Wu, Guilian Chen, Zhenkun Wen, Jing QinICCV 2021 · 55 citations
- DINO: A Conditional Energy-Based GAN for Domain TranslationKonstantinos Vougioukas, Stavros Petridis, Maja PanticICLR 2021 · 8 citations
