Chartist: Task-driven Eye Movement Control for Chart Reading
Danqing Shi, Yao Wang, Yunpeng Bai, Andreas Bulling, Antti Oulasvirta
摘要
To design data visualizations that are easy to comprehend, we need to understand how people with different interests read them. Computational models of predicting scanpaths on charts could complement empirical studies by offering estimates of user performance inexpensively; however, previous models have been limited to gaze patterns and overlooked the effects of tasks. Here, we contribute Chartist, a computational model that simulates how users move their eyes to extract information from the chart in order to perform analysis tasks, including value retrieval, filtering, and finding extremes. The novel contribution lies in a two-level hierarchical control architecture. At the high level, the model uses LLMs to comprehend the information gained so far and applies this representation to select a goal for the lower-level controllers, which, in turn, move the eyes in accordance with a sampling policy learned via reinforcement learning. The model is capable of predicting human-like task-driven scanpaths across various tasks. It can be applied in fields such as explainable AI, visualization design evaluation, and optimization. While it displays limitations in terms of generalizability and accuracy, it takes modeling in a promising direction, toward understanding human behaviors in interacting with charts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Charts-of-Thought: Enhancing LLM Visualization Literacy Through Structured Data ExtractionAmit Kumar Das, Mohammad Tarun, Klaus MuellerIEEE VIS 2025 · 被引用 6 次
- Log2Motion: Biomechanical Motion Synthesis from Touch LogsMichal Patryk Miazga, Hannah Bussmann, Antti Oulasvirta, Patrick EbelCHI 2026 · 被引用 2 次
- Making Multimodal LLMs Reliable Chart Data Extractors: A Benchmark and Training FrameworkYuchen He, Peizhi Ying, Liqi Cheng, Kuilin Peng 等CHI 2026 · 被引用 1 次
- Simulating Human Audiovisual Search BehaviorHyunsung Cho, Xuejing Luo, Byungjoo Lee, David Lindlbauer 等CHI 2026 · 被引用 1 次
它引用的顶会 Paper17
- Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied AgentsWenlong Huang, Pieter Abbeel, Deepak Pathak, Igor MordatchICML 2022 · 被引用 1,539 次
- Calliope: Automatic Visual Data Story Generation from a SpreadsheetDanqing Shi, Xinyue Xu, Fuling Sun, Yang Shi 等IEEE VIS 2020 · 被引用 179 次
- Computational Rationality as a Theory of InteractionAntti Oulasvirta, Jussi P. P. Jokinen, Andrew HowesCHI 2022 · 被引用 127 次
- Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics TasksMurtaza Dalal, Tarun Chiruvolu, Devendra Singh Chaplot, Ruslan SalakhutdinovICLR 2024 · 被引用 86 次
- Predicting Visual Importance Across Graphic Design TypesCamilo Fosco, Vincent Casser, Amish Kumar Bedi, Peter O'Donovan 等UIST 2020 · 被引用 55 次
相关 Paper
- An Adaptive Model of Gaze-based SelectionXiuli Chen, Aditya Acharya, Antti Oulasvirta, Andrew HowesCHI 2021 · 被引用 38 次
- EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided Reinforcement LearningYue Jiang, Zixin Guo, Hamed Rezazadegan Tavakoli, Luis A. Leiva 等UIST 2024 · 被引用 15 次
- ChartGaze: Enhancing Chart Understanding in LVLMs with Eye-Tracking Guided Attention RefinementAli Salamatian, Amirhossein Abaskohi, Wan-Cyuan Fan, Mir Rayat Imtiaz Hossain 等EMNLP 2025 · 被引用 1 次
- ChartSketcher: Reasoning with Multimodal Feedback and Reflection for Chart UnderstandingMuye Huang, Lingling Zhang, Jie Ma, Han Lai 等NeurIPS 2025 · 被引用 13 次
- Predicting Human Scanpaths in Visual Question AnsweringXianyu Chen, Ming Jiang, Qi ZhaoCVPR 2021
