Language Semantic Graph Guided Data-Efficient Learning
Wenxuan Ma, Shuang Li, Lincan Cai, Jingxuan Kang
Abstract
Developing generalizable models that can effectively learn from limited data and with minimal reliance on human supervision is a significant objective within the machine learning community, particularly in the era of deep neural networks. Therefore, to achieve data-efficient learning, researchers typically explore approaches that can leverage more related or unlabeled data without necessitating additional manual labeling efforts, such as Semi-Supervised Learning (SSL), Transfer Learning (TL), and Data Augmentation (DA). SSL leverages unlabeled data in the training process, while TL enables the transfer of expertise from related data distributions. DA broadens the dataset by synthesizing new data from existing examples. However, the significance of additional knowledge contained within labels has been largely overlooked in research. In this paper, we propose a novel perspective on data efficiency that involves exploiting the semantic information contained in the labels of the available data. Specifically, we introduce a Language Semantic Graph (LSG) which is constructed from labels manifest as natural language descriptions. Upon this graph, an auxiliary graph neural network is trained to extract high-level semantic relations and then used to guide the training of the primary model, enabling more adequate utilization of label knowledge. Across image, video, and audio modalities, we utilize the LSG method in both TL and SSL scenarios and illustrate its versatility in significantly enhancing performance compared to other data-efficient learning approaches. Additionally, our in-depth analysis shows that the LSG method also expedites the training process.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f47a3989-3011-4772-9931-c60efa5dd4ecCited by top-tier papers2
- Enhancing Cross-Modal Fine-Tuning with Gradually Intermediate Modality GenerationLincan Cai, Shuang Li, Wenxuan Ma, Jingxuan Kang et al.ICML 2024 · 4 citations
- Language-Assisted Debiasing and Smoothing for Foundation Model-Based Semi-Supervised LearningNa Zheng, Xuemeng Song, Xue Dong, Aashish Nikhil Ghosh et al.CVPR 2025
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
Related papers
- Enhancing Semi-Supervised Learning with Cross-Modal KnowledgeHui Zhu, Yongchun Lü, Hongbin Wang, Xunyi Zhou et al.ACM MM 2022 · 4 citations
- Efficient End-to-end Language Model Fine-tuning on GraphsRui Xue, Xipeng Shen, Ruozhou Yu, Xiaorui LiuKDD 2025 · 1 citation
- Schema-aware Reference as Prompt Improves Data-Efficient Knowledge Graph ConstructionYunzhi Yao, Shengyu Mao, Ningyu Zhang, Xiang Chen et al.SIGIR 2023 · 23 citations
- Compressing LLM Knowledge into Graph Representations for Text-attributed Graphs LearningRunhuai Chen, Dian Shen, Dandan Zhang, Kaihong Huang et al.ACL 2026
- Combining LLM Semantic Reasoning with GNN Structural Modeling for Multi-View Multi-Label Feature SelectionZhiqi Chen, Yuzhou Liu, Jiarui Liu, Wanfu GaoAAAI 2026 · 1 citation
