Inspiration Seeds: Learning Non-Literal Visual Combinations for Generative Exploration
Kfir Goldberg, Elad Richardson, Yael Vinker
摘要
While generative models have become powerful tools for image synthesis, they are typically optimized for executing carefully crafted textual prompts, offering limited support for the open-ended visual exploration that often precedes idea formation. In contrast, designers frequently draw inspiration from loosely connected visual references, seeking emergent connections that spark new ideas. We propose Inspiration Seeds, a generative framework that shifts image generation from final execution to exploratory ideation. Given two input images, our model produces diverse, visually coherent compositions that reveal latent relationships between inputs, without relying on user-specified text prompts. Our approach is feed-forward, trained on synthetic triplets of decomposed visual aspects derived entirely through visual means: we use CLIP Sparse Autoencoders to extract editing directions in CLIP latent space and isolate concept pairs. By removing the reliance on language and enabling fast, intuitive recombination, our method supports visual ideation at the early and ambiguous stages of creative work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
相关 Paper
- GenQuery: Supporting Expressive Visual Search with Generative ModelsKihoon Son, DaEun Choi, Tae Soo Kim, Young-Ho Kim 等CHI 2024 · 被引用 48 次
- CreativeConnect: Supporting Reference Recombination for Graphic Design Ideation with Generative AIDaEun Choi, Sumin Hong, Jeongeon Park, John Joon Young Chung 等CHI 2024 · 被引用 116 次
- IP-Composer: Semantic Composition of Visual ConceptsSara Dorfman, Dana Cohen-Bar, Rinon Gal, Daniel Cohen-OrSIGGRAPH 2025 · 被引用 4 次
- InspirationGraph for Progressive Design Space ExplorationSuxiang Ling, Yi Xiao, Yang Zhou, Ruoxuan Ma 等CHI 2026
- CLIP2StyleGAN: Unsupervised Extraction of StyleGAN Edit DirectionsRameen Abdal, Peihao Zhu, John Femiani, Niloy J. Mitra 等SIGGRAPH 2022 · 被引用 76 次
