Protecting multimodal large language models against misleading visualizations
Jonathan Tonglet, Tinne Tuytelaars, Marie-Francine Moens, Iryna Gurevych
Abstract
Visualizations play a pivotal role in daily communication in an increasingly data-driven world. Research on multimodal large language models (MLLMs) for automated chart understanding has accelerated massively, with steady improvements on standard benchmarks. However, for MLLMs to be reliable, they must be robust to misleading visualizations, i.e., charts that distort the underlying data, leading readers to draw inaccurate conclusions. Here, we uncover an important vulnerability: MLLM question-answering (QA) accuracy on misleading visualizations drops on average to the level of the random baseline. To address this, we provide the first comparison of six inference-time methods to improve QA performance on misleading visualizations, without compromising accuracy on non-misleading ones. We find that two methods, table-based QA and redrawing the visualization, are effective, with improvements of up to 19.6 percentage points. We make our code and data available. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext aba7b8a2-5d01-4854-8044-03e1d7b650efCited by top-tier papers3
- MisVisFix: An Interactive Dashboard for Detecting, Explaining, and Correcting Misleading Visualizations using Large Language ModelsAmit Kumar Das, Klaus MuellerIEEE VIS 2025 · 5 citations
- Is this chart lying to me? Automating the detection of misleading visualizationsJonathan Tonglet, Jan Zimny, Tinne Tuytelaars, Iryna GurevychACL 2026 · 4 citations
- Unmasking Deceptive Visuals: Benchmarking Multimodal Large Language Models on Misleading Chart Question AnsweringZixin Chen, Sicheng Song, KaShun Shum, Yanna Lin et al.EMNLP 2025 · 1 citation
Builds on16
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Surfacing Visualization MiragesAndrew M. McNutt, Gordon Kindlmann, Michael CorrellCHI 2020 · 103 citations
- Mapping the Landscape of COVID-19 Crisis VisualizationsYixuan Zhang, Yifan Sun, Lace M. K. Padilla, Sumit Barua et al.CHI 2021 · 74 citations
- CALVI: Critical Thinking Assessment for Literacy in VisualizationsLily W. Ge, Yuan Cui, Matthew KayCHI 2023 · 68 citations
- VizLinter: A Linter and Fixer Framework for Data VisualizationQing Chen, Fuling Sun, Xinyue Xu, Zui Chen et al.IEEE VIS 2021 · 60 citations
Related papers
- How Good (Or Bad) Are LLMs at Detecting Misleading Visualizations?Leo Yu-Ho Lo, Huamin QuIEEE VIS 2024 · 24 citations
- ChartR: Evaluating Reasoning Accuracy and Robustness in Chart Question AnsweringXiaojun Chen, Sixiao Luo, Ziqi Liu, Min Yang et al.CVPR 2026
- Making Multimodal LLMs Reliable Chart Data Extractors: A Benchmark and Training FrameworkYuchen He, Peizhi Ying, Liqi Cheng, Kuilin Peng et al.CHI 2026 · 1 citation
- Advancing Multimodal Large Language Models in Chart Question Answering with Visualization-Referenced Instruction TuningXingchen Zeng, Haichuan Lin, Yilin Ye, Wei ZengIEEE VIS 2024 · 23 citations
- An Empirical Evaluation of the GPT-4 Multimodal Language Model on Visualization Literacy TasksAlexander Bendeck, John T. StaskoIEEE VIS 2024 · 40 citations
