SD-DVQ: Semantic-Driven Data Visualization Query Generation from Speech Queries
Haodi Zhang, Xiaohui Tang, Yuanfeng Song
摘要
In the era of big data, effectively leveraging vast amounts of data is one of the major challenges across industries, especially in data visualization, which helps decision-makers intuitively understand data. However, many existing data visualization tools require users to have certain technical skills, posing a barrier for users without a technical background. With the widespread use of voice interfaces, particularly smart assistants, the task of converting voice queries into data visualizations, known as ''Speech-to-Vis,'' has emerged. Yet, existing methods still face issues such as speech recognition errors and insufficient model generalization. We propose SD-DVQ, a Semantic-Driven Data Visualization Query Generation from Speech Queries framework, to address these issues. Unlike previous methods that rely on ASR-based cascades or completely end-to-end systems, SD-DVQ generates semantically consistent natural language queries from speech with schema information, and then improves downstream DVQ generation through schema filtering and database-feedback-based self-correction. Experiments on the SpeechNVBench benchmark show that SD-DVQ achieves 45.02% exact-match accuracy, outperforming the strongest baseline by 5.07 percentage points.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- RGVisNet: A Hybrid Retrieval-Generation Neural Framework Towards Automatic Data Visualization GenerationYuanfeng Song, Xuefang Zhao, Raymond Chi-Wing Wong, Di JiangKDD 2022 · 被引用 28 次
- Marrying Dialogue Systems with Data Visualization: Interactive Data Visualization Generation from Natural Language ConversationsYuanfeng Song, Xuefang Zhao, Raymond Chi-Wing WongKDD 2024 · 被引用 7 次
- Towards Robustness of Text-to-Visualization Translation Against Lexical and Phrasal VariabilityJinwei Lu, Yuanfeng Song, Haodi Zhang, Chen Jason Zhang 等ICDE 2025 · 被引用 3 次
- Closing the Feedback Loop in Text2Vis: Refining Visualization with Vision-Language ModelsShengze Shi, Tao Ren, Guoliang Zhu, Guan Dong Feng 等ACM MM 2025 · 被引用 2 次
- Synthesizing Natural Language to Visualization (NL2VIS) Benchmarks from NL2SQL BenchmarksYuyu Luo, Nan Tang, Guoliang Li, Chengliang Chai 等SIGMOD 2021 · 被引用 90 次
