VisAnatomy: An SVG Chart Corpus with Fine-Grained Semantic Labels
Chen Chen, Hannah K. Bako, Peihong Yu, John Hooker, Jeffrey Joyal, Simon C. Wang, Samuel Kim, Jessica Wu, Aoxue Ding, Lara Sandeep, Alex Chen, Chayanika Sinha, Zhicheng Liu
Abstract
Chart corpora, which comprise data visualizations and their semantic labels, are crucial for advancing visualization research. However, the labels in most existing corpora are high-level (e.g., chart types), hindering their utility for broader applications in the era of AI. In this paper, we contribute VisAnatomy, a corpus containing 942 real-world SVG charts produced by over 50 tools, encompassing 40 chart types and featuring structural and stylistic design variations. Each chart is augmented with multi-level fine-grained labels on its semantic components, including each graphical element's type, role, and position, hierarchical groupings of elements, group layouts, and visual encodings. In total, VisAnatomy provides labels for more than 383k graphical elements. We demonstrate the richness of the semantic labels by comparing VisAnatomy with existing corpora. We illustrate its usefulness through four applications: semantic role inference for SVG elements, chart semantic decomposition, chart type classification, and content navigation for accessibility. Finally, we discuss research opportunities to further improve VisAnatomy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 74b2f2a5-aeb1-4856-a0e1-bf37c8c72eb1Cited by top-tier papers2
- DataWink: Reusing and Adapting SVG-Based Visualization Examples with Large Multimodal ModelsLiwenhan Xie, Yanna Lin, Can Liu, Huamin Qu et al.IEEE VIS 2025 · 3 citations
- Vector Prism: Animating Vector Graphics by Stratifying Semantic StructureJooyeol Yun, Jaegul ChooCVPR 2026 · 2 citations
Builds on13
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Composition and Configuration Patterns in Multiple-View VisualizationsXi Chen, Wei Zeng, Yanna Lin, Hayder Mahdi Al-Maneea et al.IEEE VIS 2020 · 138 citations
- VisText: A Benchmark for Semantically Rich Chart CaptioningBenny J. Tang, Angie W. Boggust, Arvind SatyanarayanACL 2023 · 44 citations
- VisEval: A Benchmark for Data Visualization in the Era of Large Language ModelsNan Chen, Yuge Zhang, Jiahang Xu, Kan Ren et al.IEEE VIS 2024 · 44 citations
Related papers
- Mystique: Deconstructing SVG Charts for Layout ReuseChen Chen, Bongshin Lee, Yunhai Wang, Yunjeong Chang et al.IEEE VIS 2023 · 10 citations
- Structure-aware Visualization RetrievalHaotian Li, Yong Wang, Aoyu Wu, Huan Wei et al.CHI 2022 · 33 citations
- ChartGalaxy: A Dataset for Infographic Chart Understanding and GenerationZhen Li, Duan Li, Yukai Guo, Xinyuan Guo et al.ICLR 2026 · 16 citations
- DIVI: Dynamically Interactive VisualizationLuke S. Snyder, Jeffrey HeerIEEE VIS 2023 · 14 citations
- Consensus and Contradictions: A Cross-Organizational Analysis of Visualization Style GuidesAlvitta OttleyCHI 2026 · 1 citation
