TransPixeler: Advancing Text-to-Video Generation with Transparency
Luozhou Wang, Yijun Li, Zhifei Chen, Jui-Hsien Wang, Zhifei Zhang, He Zhang, Zhe Lin, Ying-Cong Chen
Abstract
A statue crumbling to dust as cracks spread across its surface" "A squirrel's bushy tail flicking quickly" "A portal crackling with arcane magic as it opens" "A small explosion rapidly expanding and contracting" "A massive storm forming, with swirling clouds and thunderbolts" "A motorcycle drifting and swerving in an enchanted forest" "A white dandelion shifting as seen through a magnifying glass" "A transparent glass of water with ice cubes gently swirling" * This work was done during an internship at Adobe Research.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b6924a1e-7451-4e48-9f65-55b31161ab17Cited by top-tier papers2
- LayerFlow: A Unified Model for Layer-aware Video GenerationSihui Ji, Hao Luo, Xi Chen, Yuanpeng Tu et al.SIGGRAPH 2025 · 12 citations
- VSF: Simple, Efficient, and Effective Negative Guidance in Few-Step Image Generation Models By Value Sign FlipWenqi Guo, Shan DuICLR 2026 · 3 citations
Builds on24
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific TuningYuwei Guo, Ceyuan Yang, Anyi Rao, Zhengyang Liang et al.ICLR 2024 · 1,493 citations
- Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video GenerationJay Zhangjie Wu, Yixiao Ge, Xintao Wang, Stan Weixian Lei et al.ICCV 2023 · 1,113 citations
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled DataLihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu et al.CVPR 2024 · 847 citations
Related papers
- Dream3D: Zero-Shot Text-to-3D Synthesis Using 3D Shape Prior and Text-to-Image Diffusion ModelsJiale Xu, Xintao Wang, Weihao Cheng, Yan-Pei Cao et al.CVPR 2023
- LucidDreamer: Towards High-Fidelity Text-to-3D Generation via Interval Score MatchingYixun Liang, Xin Yang, Jiantao Lin, Haodong Li et al.CVPR 2024
- ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit AdaptationDar-Yen Chen, Hamish Tennent, Ching-Wen HsuCVPR 2024
- Align Your Latents: High-Resolution Video Synthesis with Latent Diffusion ModelsAndreas Blattmann, Robin Rombach, Huan Ling, Tim Dockhorn et al.CVPR 2023
- PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention SteeringYibin Wang, Weizhong Zhang, Jianwei Zheng, Cheng JinACM MM 2024 · 9 citations
