TAROT: Task-Adaptive Refinement of LLM-prior Graphs for Few-shot Tabular Learning
Ruxue Shi, Yili Wang, Mengnan Du, Hangting Ye, Yi Chang, Xin Wang
摘要
Few-shot tabular learning provides a cost-effective approach for real-world applications where annotation is costly and collecting sufficient samples for new tasks is difficult. Existing Traditional and LLM-based methods have demonstrated effectiveness in few-shot scenarios. However, traditional methods need additional training on unlabeled or generated data, which incur significant computational overhead. In addition, LLM-based methods that directly feed raw tabular data into LLMs raise privacy and compliance concerns. More importantly, both paradigms largely overlook the semantic relationships between features, which provide structural and semantic prior for constructing a semantic graph. Semantic graph is essential for modeling meaningful feature interactions in few-shot scenarios. In this paper, we propose TAROT, a GNN-based framework that encodes the structural and semantic prior by constructing and refining a task-adaptive semantic graph from this prior, thereby improving predictive performance in few-shot tabular learning. TAROT first encodes heterogeneous tabular data into unified node semantic representations via a Unified Semantic Tabular Node Encoder (USTNE). Then, it prompts LLMs to infer the semantic relationship between features based on the task description and feature names to construct a semantic graph. To mitigate structural noise introduced by the hallucination of LLMs, TAROT introduces Task-adaptive Semantic Graph Refinement that prunes spurious or task-unrelated edges and adds missing task-related ones, aligning the graph structure with the downstream objective. Finally, a GNN performs message passing over the refined graph to capture task-related semantic dependencies for prediction. Extensive experiments on various few-shot tabular learning benchmarks demonstrate the superior performance of TAROT, establishing it as a state-of-the-art approach in this domain.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Extracting Training Data from Large Language ModelsNicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski 等USENIX Security 2021 · 被引用 2,866 次
- Structure-Augmented Text Representation Learning for Efficient Knowledge Graph CompletionBo Wang, Tao Shen, Guodong Long, Tianyi Zhou 等WWW 2021 · 被引用 322 次
- Well-tuned Simple Nets Excel on Tabular DatasetsArlind Kadra, Marius Lindauer, Frank Hutter, Josif GrabockaNeurIPS 2021 · 被引用 288 次
- Scarf: Self-Supervised Contrastive Learning using Random Feature CorruptionDara Bahri, Heinrich Jiang, Yi Tay, Donald MetzlerICLR 2022 · 被引用 233 次
- Few-Shot Image Recognition With Knowledge TransferZhimao Peng, Zechao Li, Junge Zhang, Yan Li 等ICCV 2019 · 被引用 230 次
相关 Paper
- HeGTa: Leveraging Heterogeneous Graph-enhanced Large Language Models for Few-shot Complex Table UnderstandingRihui Jin, Yu Li, Guilin Qi, Nan Hu 等AAAI 2025 · 被引用 1 次
- STUNT: Few-shot Tabular Learning with Self-generated Tasks from Unlabeled TablesJaehyun Nam, Jihoon Tack, Kyungmin Lee, Hankook Lee 等ICLR 2023 · 被引用 2 次
- Large Language Models Can Automatically Engineer Features for Few-Shot Tabular LearningSungwon Han, Jinsung Yoon, Sercan Ö. Arik, Tomas PfisterICML 2024 · 被引用 81 次
- D2R2: Diffusion-based Representation with Random Distance Matching for Tabular Few-shot LearningRuoxue Liu, Linjiajie Fang, Wenjia Wang, Bingyi JingNeurIPS 2024 · 被引用 7 次
- Language Model Representations for Efficient Few-Shot Tabular ClassificationInwon Kang, Parikshit Ram, Yi Zhou, Horst Samulowitz 等WWW 2026
