Revisiting Single-Table Retrieval: An Open Problem Under 360° Stress Tests
Chenyu Yang, Ziyu Jiang, Junhao Li, Yuyu Luo, Ju Fan, Nan Tang
摘要
Single-table retrieval (STR)-selecting the most relevant table from a data lake to answer a natural language question-remains an open challenge despite recent progress in dense retrieval. A conclusive assessment remains out of reach because (i) publicly available datasets are limited to fully reflect real-world complexity, (ii) experimental pipelines lack standardization, preventing fair comparison, and (iii) evaluations focus narrowly on standard performance on a fixed data set, overlooking robustness, generalization, and near-duplicate handling. We conduct the first systematic, large-scale investigation of dense retrieval for STR and contribute four advances: (1) a unified pipeline that factorizes STR into four design spaces: encoder architecture, model training, table structure encoding, and linearization; (2) TR360, a modular testbed that instantiates this pipeline and enables multi-angle stress testing under realistic, evolving data-lake scenarios; (3) TRBench, a benchmark that fuses six heterogeneous corpora, supplies realistic questions with varying reasoning depth, and includes a synthetic-data generator for rapid domain adaptation; and (4) the largest empirical study to date, spanning over 10 models, which distills actionable guidelines while revealing persistent weaknesses in generalization, robustness, and near-duplicate discrimination. Our results demonstrate that, despite notable advances, singletable retrieval remains unsolved, with persistent weaknesses in generalization, robustness, and the discrimination of nearduplicate tables. All code, data, and evaluation scripts are at https://github.com/CTXY/table_retrieval.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- T2R-BENCH: A Benchmark for Real World Table-to-Report TaskJie Zhang, Changzai Pan, Sishi Xiong, Kaiwen Wei 等EMNLP 2025 · 被引用 2 次
- REaR : Retrieve, Expand and Refine for Effective Multitable RetrievalRishita Agarwal, Himanshu Singhal, Peter Baile Chen, Manan Roy Choudhury 等ACL 2026 · 被引用 2 次
- How Far Can LLM Agents Reason with Tables? Benchmarking Multi-Turn Agentic Table Question Answering in the WildJingwang Huang, Jie Zhang, Haoyang Zeng, Changzai Pan 等ICML 2026
- CRAFT: Training-Free Cascaded Retrieval for Tabular QAAdarsh Singh, Kushal Raj Bhandari, Jianxi Gao, Soham Dan 等ACL 2026 · 被引用 2 次
- SURE or Not? Investigating Semantic Understanding in Dense Retrieval ModelsLingdi Kong, Xuanang Chen, Ben He, Le SunACL 2026
