Let the Chart Spark: Embedding Semantic Context into Chart with Text-to-Image Generative Model
Shishi Xiao, Suizi Huang, Yue Lin, Yilin Ye, Wei Zeng
Abstract
Pictorial visualization seamlessly integrates data and semantic context into visual representation, conveying complex information in an engaging and informative manner. Extensive studies have been devoted to developing authoring tools to simplify the creation of pictorial visualizations. However, mainstream works follow a retrieving-and-editing pipeline that heavily relies on retrieved visual elements from a dedicated corpus, which often compromise data integrity. Text-guided generation methods are emerging, but may have limited applicability due to their predefined entities. In this work, we propose ChartSpark, a novel system that embeds semantic context into chart based on text-to-image generative models. ChartSpark generates pictorial visualizations conditioned on both semantic context conveyed in textual inputs and data information embedded in plain charts. The method is generic for both foreground and background pictorial generation, satisfying the design practices identified from empirical research into existing pictorial visualizations. We further develop an interactive visual interface that integrates a text analyzer, editing module, and evaluation module to enable users to generate, modify, and assess pictorial visualizations. We experimentally demonstrate the usability of our tool, and conclude with a discussion of the potential of using text-to-image generative models combined with an interactive interface for visualization design.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 19bed08d-ef91-40a1-8fab-03ea0e7036e5Cited by top-tier papers19
- PlantoGraphy: Incorporating Iterative Design Process into Generative Artificial Intelligence for Landscape RenderingRong Huang, Haichuan Lin, Chuanzhang Chen, Kang Zhang et al.CHI 2024 · 45 citations
- TypeDance: Creating Semantic Typographic Logos from Image through Personalized GenerationShishi Xiao, Liangwei Wang, Xiaojuan Ma, Wei ZengCHI 2024 · 35 citations
- SketchFlex: Facilitating Spatial-Semantic Coherence in Text-to-Image Generation with Region-Based SketchesHaichuan Lin, Yilin Ye, Jiazhi Xia, Wei ZengCHI 2025 · 25 citations
- Advancing Multimodal Large Language Models in Chart Question Answering with Visualization-Referenced Instruction TuningXingchen Zeng, Haichuan Lin, Yilin Ye, Wei ZengIEEE VIS 2024 · 23 citations
- Beyond Numbers: Creating Analogies to Enhance Data Comprehension and Communication with Generative AIQing Chen, Wei Shuai, Jiyao Zhang, Zhida Sun et al.CHI 2024 · 18 citations
Builds on25
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- Supporting Expressive and Faithful Pictorial Visualization Design with Visual Style TransferYang Shi, Pei Liu, Siji Chen, Mengdi Sun et al.IEEE VIS 2022 · 40 citations
- Semantic-Structural Alignment for Generative Pictorial ChartsZhida Sun, Yulin Zhang, Zheng Gu, Min Lu et al.SIGGRAPH 2026
- ChArtist: Generating Pictorial Charts with Unified Spatial and Subject ControlShishi Xiao, Tongyu Zhou, David H. Laidlaw, Gromit Yeuk-Yin ChanCVPR 2026 · 2 citations
- Text2Vis: A Challenging and Diverse Benchmark for Generating Multimodal Visualizations from TextMizanur Rahman, Md. Tahmid Rahman Laskar, Shafiq Joty, Enamul HoqueEMNLP 2025 · 1 citation
- StructGraphics: Flexible Visualization Design through Data-Agnostic and Reusable Graphical StructuresTheophanis TsandilasIEEE VIS 2020 · 26 citations
