Infogen: Generating Complex Statistical Infographics from Documents
Akash Ghosh, Aparna Garimella, Pritika Ramu, Sambaran Bandyopadhyay, Sriparna Saha
摘要
Statistical infographics are powerful tools that simplify complex data into visually engaging and easy-to-understand formats. Despite advancements in AI, particularly with LLMs, existing efforts have been limited to generating simple charts, with no prior work addressing the creation of complex infographics from textheavy documents that demand a deep understanding of the content. We address this gap by introducing the task of generating statistical infographics composed of multiple subcharts (e.g., line, bar, pie) that are contextually accurate, insightful, and visually aligned. To achieve this, we define infographic metadata, that includes its title and textual insights, along with sub-chart-specific details such as their corresponding data, alignment, etc. We also present Infodat, the first benchmark dataset for text-to-infographic metadata generation, where each sample links a document to its metadata. We propose Infogen, a two-stage framework where fine-tuned LLMs first generate metadata, which is then converted into infographic code. Extensive evaluations on Infodat demonstrate that Infogen achieves state-of-the-art performance, outperforming both closed and opensource LLMs in text-to-statistical infographic generation. The sample datapoints from Infodat can be accessed through this link * Work done during internship at Adobe Research. 1 While the term infographic can refer to a wide range of visual illustrations, we restrict ourselves to only those visuals
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- CLINIC : Evaluating Multilingual Trustworthiness in Language Models for HealthcareAkash Ghosh, Srivarshinee Sridhar, Raghav Kaushik Ravi, Muhsin Muhsin 等ICML 2026 · 被引用 6 次
- Beyond Static Artifacts: An Evolutionary Framework for Synthetic Claim GenerationYeqing Teng, Jiasheng Si, Shuxia Lin, Linhai Zhang 等ACL 2026
它引用的顶会 Paper10
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- QLoRA: Efficient Finetuning of Quantized LLMsTim Dettmers, Artidoro Pagnoni, Ari Holtzman, Luke ZettlemoyerNeurIPS 2023 · 被引用 5,863 次
- ReFT: Representation Finetuning for Language ModelsZhengxuan Wu, Aryaman Arora, Zheng Wang, Atticus Geiger 等NeurIPS 2024 · 被引用 233 次
- Calliope: Automatic Visual Data Story Generation from a SpreadsheetDanqing Shi, Xinyue Xu, Fuling Sun, Yang Shi 等IEEE VIS 2020 · 被引用 179 次
- UniChart: A Universal Vision-language Pretrained Model for Chart Comprehension and ReasoningAhmed Masry, Parsa Kavehzadeh, Do Xuan Long, Enamul Hoque 等EMNLP 2023 · 被引用 48 次
相关 Paper
- IGenBench: Benchmarking the Reliability of Text-to-Infographic GenerationYinghao Tang, Xueding Liu, Boyuan Zhang, Tingfeng Lan 等ACL 2026 · 被引用 9 次
- ChartGalaxy: A Dataset for Infographic Chart Understanding and GenerationZhen Li, Duan Li, Yukai Guo, Xinyuan Guo 等ICLR 2026 · 被引用 16 次
- InfoDet: A Dataset for Infographic Element DetectionJiangning Zhu, Yuxing Zhou, Zheng Wang, Juntao Yao 等ICLR 2026 · 被引用 4 次
- Doc2Chart: Intent-Driven Zero-Shot Chart Generation from DocumentsAkriti Jain, Pritika Ramu, Aparna Garimella, Apoorv SaxenaEMNLP 2025
- Retrieve-Then-Adapt: Example-based Automatic Generation for Proportion-related InfographicsChunyao Qian, Shizhao Sun, Weiwei Cui, Jian-Guang Lou 等IEEE VIS 2020 · 被引用 53 次
