Lune

CVPR2025顶会

Compass Control: Multi Object Orientation Control for Text-to-Image Generation

Rishubh Parihar, Vaibhav Agrawal, Sachidanand VS, Venkatesh Babu Radhakrishnan

2025年份
6顶会引用

摘要

Personalization with 3D Orientation Control Few unposed Input Images 'A photo of V* car in front of the leaning tower of Pisa in Italy' 0.523 1.047 2.617 3.665 3.141 0.0 1.047 2.094 5.235 3.123 'A photo of a mother walking with a pram on a snowy street, festive Christmas lights, beautiful winter evening scene' 0.675, 0.725 0.80, 0.60 * equal contribution. † work done during an internship at VAL, IISc tion. In this work, we address the problem of multi-object orientation control in text-to-image diffusion models. This enables the generation of diverse multi-object scenes with precise orientation control for each object. The key idea is to condition the diffusion model with a set of orientationaware compass tokens, one for each object, along with text tokens. A light-weight encoder network predicts these com-This CVPR paper is the Open Access version, provided by the Computer Vision Foundation. Except for this watermark, it is identical to the accepted version; the final published version of the proceedings is available on IEEE Xplore.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper6

问问它们各自怎么用它

它引用的顶会 Paper41

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖