Lune

CVPR2025顶会

LaTexBlend: Scaling Multi-concept Customized Generation with Latent Textual Blending

Jian Jin, Zhenbo Yu, Yang Shen, Zhenyong Fu, Jian Yang

2025年份
4顶会引用

摘要

tioned after the text encoder and a linear projection. LA-TEXBLEND customizes each concept individually, storing them in a concept bank with a compact representation of latent textual features that captures sufficient concept information to ensure high fidelity. At inference, concepts from the bank can be freely and seamlessly combined in the latent textual space, offering two key merits for multiconcept generation: 1) excellent scalability, and 2) significant reduction of denoising deviation, preserving coherent layouts. Extensive experiments demonstrate that LATEXBLEND can flexibly integrate multiple customized concepts with harmonious structures and high subject fidelity, substantially outperforming baselines in both generation quality and computational efficiency. Project page: https://jinjianrick.github.io/latexblend/ This CVPR paper is the Open Access version, provided by the Computer Vision Foundation. Except for this watermark, it is identical to the accepted version; the final published version of the proceedings is available on IEEE Xplore. MuDI Concept bank OMG Mix-of-Show V 2 * bear plushie sitting on V 1 * chair. V 7 * dog playing V 8 * guitar, surrounded by V 6 * flower, with V 10 * lighthouse in the background. V 11 * cat sitting next to V 12 * teddybear, with V 6 * flower blooming beside them, with V 10 * lighthouse and V 9 * barn in the background. Two kids wearing V 3 * jacket and V 4 * shoes, playing with V 5 * dog.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext b449df33-c50b-4425-bc3c-e1e152113f6b

引用它的顶会 Paper4

问问它们各自怎么用它

它引用的顶会 Paper35

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖