Learn from Yesterday: A Semi-supervised Continual Learning Method for Supervision-Limited Text-to-SQL Task Streams
Yongrui Chen, Xinnan Guo, Tongtong Wu, Guilin Qi, Yang Li, Yang Dong
摘要
Conventional text-to-SQL studies are limited to a single task with a fixed-size training and test set. When confronted with a stream of tasks common in real-world applications, existing methods struggle with the problems of insufficient supervised data and high retraining costs. The former tends to cause overfitting on unseen databases for the new task, while the latter makes a full review of instances from past tasks impractical for the model, resulting in forgetting of learned SQL structures and database schemas. To address the problems, this paper proposes integrating semi-supervised learning (SSL) and continual learning (CL) in a stream of text-to-SQL tasks and offers two promising solutions in turn. The first solution Vanilla is to perform self-training, augmenting the supervised training data with predicted pseudo-labeled instances of the current task, while replacing the full volume retraining with episodic memory replay to balance the training efficiency with the performance of previous tasks. The improved solution SFNet takes advantage of the intrinsic connection between CL and SSL. It uses in-memory past information to help current SSL, while adding high-quality pseudo instances in memory to improve future replay. The experiments on two datasets shows that SFNet outperforms the widely-used SSL-only and CL-only baselines on multiple metrics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Parameterizing Context: Unleashing the Power of Parameter-Efficient Fine-Tuning and In-Context Tuning for Continual Table Semantic ParsingYongrui Chen, Shenyu Zhang, Guilin Qi, Xinnan GuoNeurIPS 2023 · 被引用 11 次
- Filling Memory Gaps: Enhancing Continual Semantic Parsing via SQL Syntax Variance-Guided LLMs Without Real Data ReplayRuiheng Liu, Jinyu Zhang, Yanqi Song, Yu Zhang 等AAAI 2025 · 被引用 5 次
- K-DeCore: Facilitating Knowledge Transfer in Continual Structured Knowledge Reasoning via Knowledge DecouplingYongrui Chen, Yi Huang, Yunchang Liu, Shenyu Zhang 等NeurIPS 2025 · 被引用 2 次
- Self-Reinforcing Prototype Evolution with Dual-Knowledge Cooperation for Semi-Supervised Lifelong Person Re-IdentificationKunlun Xu, Fan Zhuo, Jiangmeng Li, Xu Zou 等ICCV 2025 · 被引用 2 次
- Exploiting Presentative Feature Distributions for Parameter-Efficient Continual Learning of Large Language ModelsXin Cheng, Jiabo Ye, Haiyang Xu, Ming Yan 等ICML 2025
它引用的顶会 Paper9
- Text-to-SQL Generation for Question Answering on Electronic Medical RecordsPing Wang, Tian Shi, Chandan K. ReddyWWW 2020 · 被引用 148 次
- Continual Relation Learning via Episodic Memory Activation and ReconsolidationXu Han, Yi Dai, Tianyu Gao, Yankai Lin 等ACL 2020 · 被引用 92 次
- Pretrained Language Model in Continual Learning: A Comparative StudyTongtong Wu, Massimo Caccia, Zhuang Li, Yuan-Fang Li 等ICLR 2022 · 被引用 76 次
- GraPPa: Grammar-Augmented Pre-Training for Table Semantic ParsingTao Yu, Chien-Sheng Wu, Xi Victoria Lin, Bailin Wang 等ICLR 2021 · 被引用 59 次
- RAT-SQL: Relation-Aware Schema Encoding and Linking for Text-to-SQL ParsersBailin Wang, Richard Shin, Xiaodong Liu, Oleksandr Polozov 等ACL 2020 · 被引用 39 次
相关 Paper
- Looking Back on Learned Experiences For Class/task Incremental LearningMozhgan PourKeshavarz, Guoying Zhao, Mohammad SabokrouICLR 2022 · 被引用 42 次
- Beyond Buffer Limits: Energy-Based Data Reassembly for Continual LearningZhenyi Wang, Yixuan Sun, Yue Wang, Zhong Chen 等ICML 2026
- Label Delay in Online Continual LearningBotos Csaba, Wenxuan Zhang, Matthias Müller, Ser Nam Lim 等NeurIPS 2024 · 被引用 12 次
- Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized RehearsalJianheng Huang, Leyang Cui, Ante Wang, Chengyi Yang 等ACL 2024 · 被引用 13 次
- Semi-supervised Drifted Stream Learning with Short LookbackWeijieying Ren, Pengyang Wang, Xiaolin Li, Charles E. Hughes 等KDD 2022 · 被引用 10 次
