Personalized Visual Content Generation in Conversational Systems
Xianquan Wang, Zhaocheng Du, Huibo Xu, Shukang Yin, Yupeng Han, Jieming Zhu, Kai Zhang, Qi Liu
摘要
With the rapid progress of large language models (LLMs) and diffusion models, there has been growing interest in personalized content generation. However, current conversational systems often present the same recommended content to all users, falling into the dilemma of "one-size-fits-all." To break this limitation and boost user engagement, in this paper, we introduce PCG ( P ersonalized Visual C ontent G eneration), a unified framework for personalizing item images within conversational systems. We tackle two key bottlenecks: the depth of personalization and the fidelity of generated images. Specifially, an LLM-powered Inclinations Analyzer is adopted to capture user likes and dislikes from context to construct personalized prompts. Moreover, we design a dual-stage LoRA mechanism— Global LoRA for understanding task-specific visual style, and Local LoRA for capturing preferred visual elements from conversation history. During training, we introduce the visual content condition method to ensure LoRA learns both historical visual context and maintains fidelity to the original item images. Extensive experiments on benchmark conversational datasets—including objective metrics and GPT-based evaluations—demonstrate that our framework outperforms strong baselines, which highlight its potential to redefine personalization in visual content generation for conversational scenarios like e-commerce and real-world recommendation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- AccKV: Towards Efficient Audio-Video LLMs Inference via Adaptive-Focusing and Cross-Calibration KV Cache OptimizationZhonghua Jiang, Kui Chen, Kunxi Li, Keting Yin 等AAAI 2026 · 被引用 4 次
- MemWeaver: A Hierarchical Memory from Textual Interactive Behaviors for Personalized GenerationShuo Yu, Mingyue Cheng, Daoyu Wang, Qi Liu 等WWW 2026 · 被引用 3 次
- Length-Adaptive Interest Network for Balancing Long and Short Sequence Modeling in CTR PredictionZhicheng Zhang, Zhaocheng Du, Jieming Zhu, Jiwei Tang 等AAAI 2026 · 被引用 2 次
它引用的顶会 Paper13
- An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual InversionRinon Gal, Yuval Alaluf, Yuval Atzmon, Or Patashnik 等ICLR 2023 · 被引用 464 次
- Towards Unified Conversational Recommender Systems via Knowledge-Enhanced Prompt LearningXiaolei Wang, Kun Zhou, Ji-Rong Wen, Wayne Xin ZhaoKDD 2022 · 被引用 143 次
- Unified Language-Vision Pretraining in LLM with Dynamic Discrete Visual TokenizationYang Jin, Kun Xu, Liwei Chen, Chao Liao 等ICLR 2024 · 被引用 87 次
- Learning to Rewrite Prompts for Personalized Text GenerationCheng Li, Mingyang Zhang, Qiaozhu Mei, Weize Kong 等WWW 2024 · 被引用 54 次
- PMG : Personalized Multimodal Generation with Large Language ModelsXiaoteng Shen, Rui Zhang, Xiaoyan Zhao, Jieming Zhu 等WWW 2024 · 被引用 40 次
相关 Paper
- ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference AlignmentZhipeng Bian, Jieming Zhu, Qijiong Liu, Wang Lin 等EMNLP 2025
- ImageGem: In-the-wild Generative Image Interaction Dataset for Generative Model PersonalizationYuanhe Guo, Linxi Xie, Zhuoran Chen, Kangrui Yu 等ICCV 2025
- VisualLens: Personalization through Task-Agnostic Visual HistoryWang Bill Zhu, Deqing Fu, Kai Sun, Yi Lu 等NeurIPS 2025
- CRAFT-LoRA: Content-Style Personalization via Rank-Constrained Adaptation and Training-Free FusionYu Li, Yujun Cai, Chi ZhangCVPR 2026 · 被引用 2 次
- Personalize Your Gaussian: Consistent 3D Scene Personalization from a Single ImageYuxuan Wang, Xuanyu Yi, Qingshan Xu, Yuan Zhou 等AAAI 2026 · 被引用 2 次
