DS²-ABSA: Dual-Stream Data Synthesis with Label Refinement for Few-Shot Aspect-Based Sentiment Analysis
Hongling Xu, Yice Zhang, Qianlong Wang, Ruifeng Xu
Abstract
Recently developed large language models (LLMs) have presented promising new avenues to address data scarcity in low-resource scenarios. In few-shot aspect-based sentiment analysis (ABSA), previous efforts have explored data augmentation techniques, which prompt LLMs to generate new samples by modifying existing ones. However, these methods fail to produce adequately diverse data, impairing their effectiveness. Additionally, some studies apply incontext learning for ABSA by using specific instructions and a few selected examples as prompts. Though promising, LLMs often yield labels that deviate from task requirements. To overcome these limitations, we propose DS 2 -ABSA, a dual-stream data synthesis framework targeted for few-shot ABSA. It leverages LLMs to synthesize data from two complementary perspectives: key-point-driven and instancedriven, which effectively generate diverse and high-quality ABSA samples in low-resource settings. Furthermore, a label refinement module is integrated to improve the synthetic labels. Extensive experiments demonstrate that DS 2 -ABSA significantly outperforms previous fewshot ABSA solutions and other LLM-oriented data generation methods. Our code and synthetic data are available at https://github. com/HITSZ-HLT/DS2-ABSA .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on11
- MetaMath: Bootstrap Your Own Mathematical Questions for Large Language ModelsLonghui Yu, Weisen Jiang, Han Shi, Jincheng Yu et al.ICLR 2024 · 637 citations
- Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP TasksYizhong Wang, Swaroop Mishra, Pegah Alipoormolabashi, Yeganeh Kordi et al.EMNLP 2022 · 238 citations
- DAGA: Data Augmentation with a Generation Approach forLow-resource Tagging TasksBosheng Ding, Linlin Liu, Lidong Bing, Canasai Kruengkrai et al.EMNLP 2020 · 132 citations
- UniversalNER: Targeted Distillation from Large Language Models for Open Named Entity RecognitionWenxuan Zhou, Sheng Zhang, Yu Gu, Muhao Chen et al.ICLR 2024 · 118 citations
- MELM: Data Augmentation with Masked Entity Language Modeling for Low-Resource NERRan Zhou, Xin Li, Ruidan He, Lidong Bing et al.ACL 2022 · 114 citations
Related papers
- LACA: Improving Cross-lingual Aspect-Based Sentiment Analysis with LLM Data AugmentationJakub Smíd, Pavel Pribán, Pavel KrálACL 2025 · 5 citations
- Cross-Domain Data Augmentation with Domain-Adaptive Language Modeling for Aspect-Based Sentiment AnalysisJianfei Yu, Qiankun Zhao, Rui XiaACL 2023 · 19 citations
- DimABSA: Building Multilingual and Multidomain Datasets for Dimensional Aspect-Based Sentiment AnalysisLung-Hao Lee, Liang-Chih Yu, Natalia V. Loukachevitch, Ilseyar Alimova et al.ACL 2026 · 1 citation
- CLAOCS-TX: Cross-Lingual Triplet Extraction with Aspect-Opinion-Aware Code-Switched Prompting and LLM-Guided Contrastive DistillationLipika Dewangan, Chandresh Kumar MauryaACL 2026
- VERO: Verification and Zero-Shot Feedback Acquisition for Few-Shot Multimodal Aspect-Level Sentiment ClassificationKai Sun, Hao Wu, Bin Shi, Samuel Mensah et al.AAAI 2025 · 1 citation
