CubeComposer: Spatio-Temporal Autoregressive 4K 360deg Video Generation from Perspective Video
Lingen Li, Guangzhi Wang, Xiaoyu Li, Zhaoyang Zhang, Qi Dou, Jinwei Gu, Tianfan Xue, Ying Shan
2026Year
10Citations
Abstract
Frame 1 Frame 25 Frame 1 Frame 25 Frame 1 Frame 25 Frame 1 Frame 25 Frame 20 Frame 15 Frame 10 Frame 5 Frame 1 Frame 25 Frame 20 Frame 15 Frame 10 Frame 5 Project Argus Generated 360¡ Video (1024×512) Argus + VEnhancer Super-resolved (SR) 360¡ Video (2048×1024) CubeComposer (Ours) Generated 360¡ Video (3840×1920, native 4K w/o SR) Frame 1 Frame 25 Frame 20 Frame 15
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5da57396-3ccf-4e32-9826-2c7c1a1e7388Builds on33
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 3,959 citations
Related papers
- UltraGen: High-Resolution Video Generation with Hierarchical AttentionTeng Hu, Jiangning Zhang, Zihan Su, Ran YiAAAI 2026 · 7 citations
- Argus: Bandwidth-Efficient Live Multiview Video Streaming via Sparse-View Gaussian ReconstructionYizong Wang, Hongbo Ning, Haohua Wang, Yutao Yuan et al.INFOCOM 2026
- Beyond Perspective: Neural 360-Degree Video CompressionAndy Regensky, Marc Windsheimer, Fabian Brand, André KaupICCV 2025 · 2 citations
- Argus: Real-Time HQ Video Decoding with CPU Coordinating on Consumer DevicesQiang Chen, Changlong LiRTSS 2024 · 1 citation
- VR Viewport Pose Model for Quantifying and Exploiting Frame CorrelationsYing Chen, Hojung Kwon, Hazer Inaltekin, Maria GorlatovaINFOCOM 2022 · 10 citations
