Lune

CVPR2023顶会

Unite and Conquer: Plug & Play Multi-Modal Synthesis Using Diffusion Models

Nithin Gopalakrishnan Nair, Wele Gedara Chaminda Bandara, Vishal M. Patel

2023年份
6顶会引用

摘要

Tibetan terrier Teddy Bear Triceratops Tree frog Otterhound Tibetan terrier Teddy Bear Triceratops Tree frog Otterhound GLIDE [21] OURS (b) (Face, hair) semantic labels, Text-→ Facial image (c) Sketch, Text-→ Facial image TediGAN [45] OURS Semantic Label This person is chubby and has wavy black hair An old person with brown hair This person has black hair and wears beard This person has brown hair and wears eyeglasses Sketch This person has blonde hair and black eyebrows This person has brown hair and dark skin tone strategy. We also introduce a novel reliability parameter that allows using different off-the-shelf diffusion models trained across various datasets during sampling time alone to guide it to the desired outcome satisfying multiple constraints. We perform experiments on various standard multimodal tasks to demonstrate the effectiveness of our approach. More details can be found at: https://nithin- gk.github.io/projectpages/Multidiff

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper6

问问它们各自怎么用它

它引用的顶会 Paper18

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖