Hierarchical Patch VAE-GAN: Generating Diverse Videos from a Single Sample
Shir Gur, Sagie Benaim, Lior Wolf
摘要
We consider the task of generating diverse and novel videos from a single video sample. Recently, new hierarchical patch-GAN based approaches were proposed for generating diverse images, given only a single sample at training time. Moving to videos, these approaches fail to generate diverse samples, and often collapse into generating samples similar to the training video. We introduce a novel patchbased variational autoencoder (VAE) which allows for a much greater diversity in generation. Using this tool, a new hierarchical video generation scheme is constructed: at coarse scales, our patch-VAE is employed, ensuring samples are of high diversity. Subsequently, at finer scales, a patch-GAN renders the fine details, resulting in high quality videos. Our experiments show that the proposed method produces diverse samples in both the image domain, and the more challenging video domain. Our code and supplementary material (SM) with additional samples are available at https://shirgur.github.io/hp-vae-gan .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video GenerationJay Zhangjie Wu, Yixiao Ge, Xintao Wang, Stan Weixian Lei 等ICCV 2023 · 被引用 1,113 次
- StyleGAN-V: A Continuous Video Generator with the Price, Image Quality and Perks of StyleGAN2Ivan Skorokhodov, Sergey Tulyakov, Mohamed ElhoseinyCVPR 2022 · 被引用 167 次
- SinDDM: A Single Image Denoising Diffusion ModelVladimir Kulikov, Shahar Yadin, Matan Kleiner, Tomer MichaeliICML 2023 · 被引用 113 次
- A Hierarchical Transformation-Discriminating Generative Model for Few Shot Anomaly DetectionShelly Sheynin, Sagie Benaim, Lior WolfICCV 2021 · 被引用 106 次
- Drop the GAN: In Defense of Patches Nearest Neighbors as Single Image Generative ModelsNiv Granot, Ben Feinstein, Assaf Shocher, Shai Bagon 等CVPR 2022 · 被引用 60 次
它引用的顶会 Paper3
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 被引用 933 次
- PatchVAE: Learning Local Latent Codes for RecognitionKamal Gupta, Saurabh Singh, Abhinav ShrivastavaCVPR 2020
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten 等CVPR 2020
相关 Paper
- Generating Diverse Structure for Image Inpainting With Hierarchical VQ-VAEJialun Peng, Dong Liu, Songcen Xu, Houqiang LiCVPR 2021
- Deep Hierarchical Video CompressionMing Lu, Zhihao Duan, Fengqing Zhu, Zhan MaAAAI 2024 · 被引用 19 次
- Bit Prioritization in Variational Autoencoders via Progressive CodingRui Shu, Stefano ErmonICML 2022 · 被引用 9 次
- Diverse Video Generation using a Gaussian Process TriggerGaurav Shrivastava, Abhinav ShrivastavaICLR 2021 · 被引用 22 次
- Contextually Plausible and Diverse 3D Human Motion PredictionSadegh Aliakbarian, Fatemeh Sadat Saleh, Lars Petersson, Stephen Gould 等ICCV 2021 · 被引用 44 次
