Evolutionary Neural Architecture Search for Transformer in Knowledge Tracing
Shangshang Yang, Xiaoshan Yu, Ye Tian, Xueming Yan, Haiping Ma, Xingyi Zhang
摘要
Knowledge tracing (KT) aims to trace students' knowledge states by predicting whether students answer correctly on exercises. Despite the excellent performance of existing Transformer-based KT approaches, they are criticized for the manually selected input features for fusion and the defect of single global context modelling to directly capture students' forgetting behavior in KT, when the related records are distant from the current record in terms of time. To address the issues, this paper first considers adding convolution operations to the Transformer to enhance its local context modelling ability used for students' forgetting behavior, then proposes an evolutionary neural architecture search approach to automate the input feature selection and automatically determine where to apply which operation for achieving the balancing of the local/global context modelling. In the search space, the original global path containing the attention module in Transformer is replaced with the sum of a global path and a local path that could contain different convolutions, and the selection of input features is also considered. To search the best architecture, we employ an effective evolutionary algorithm to explore the search space and also suggest a search space reduction strategy to accelerate the convergence of the algorithm. Experimental results on the two largest and most challenging education datasets demonstrate the effectiveness of the architecture found by the proposed approach.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Enhancing Cognitive Diagnosis Using Un-interacted Exercises: A Collaboration-Aware Mixed Sampling ApproachHaiping Ma, Changqian Wang, Hengshu Zhu, Shangshang Yang 等AAAI 2024 · 被引用 22 次
- DisenGCD: A Meta Multigraph-assisted Disentangled Graph Learning Framework for Cognitive DiagnosisShangshang Yang, Mingyang Chen, Ziwen Wang, Xiaoshan Yu 等NeurIPS 2024 · 被引用 17 次
- RIGL: A Unified Reciprocal Approach for Tracing the Independent and Group Learning ProcessesXiaoshan Yu, Chuan Qin, Dazhong Shen, Shangshang Yang 等KDD 2024 · 被引用 11 次
- ORCDF: An Oversmoothing-Resistant Cognitive Diagnosis Framework for Student Learning in Online Education SystemsHong Qian, Shuo Liu, Mingjia Li, Bingdong Li 等KDD 2024 · 被引用 11 次
- Towards Accurate and Fair Cognitive Diagnosis via Monotonic Data AugmentationZheng Zhang, Wei Song, Qi Liu, Qingyang Mao 等NeurIPS 2024 · 被引用 10 次
它引用的顶会 Paper8
- On Layer Normalization in the Transformer ArchitectureRuibin Xiong, Yunchang Yang, Di He, Kai Zheng 等ICML 2020 · 被引用 1,388 次
- AutoFormer: Searching Transformers for Visual RecognitionMinghao Chen, Houwen Peng, Jianlong Fu, Haibin LingICCV 2021 · 被引用 335 次
- HAT: Hardware-Aware Transformers for Efficient Natural Language ProcessingHanrui Wang, Zhanghao Wu, Zhijian Liu, Han Cai 等ACL 2020 · 被引用 215 次
- GLiT: Neural Architecture Search for Global and Local Image TransformerBoyu Chen, Peixia Li, Chuming Li, Baopu Li 等ICCV 2021 · 被引用 100 次
- TextNAS: A Neural Architecture Search Space Tailored for Text RepresentationYujing Wang, Yaming Yang, Yiren Chen, Jing Bai 等AAAI 2020 · 被引用 66 次
相关 Paper
- HRKT: Hierarchical Recurrent Knowledge Tracing for Efficient Transformer-Based Long-Sequence ModelingJu-Yeong Park, Tae-Gwon Lee, Ji-Hoon BaeKDD 2026
- Tracing Knowledge Instead of Patterns: Stable Knowledge Tracing with Diagnostic TransformerYu Yin, Le Dai, Zhenya Huang, Shuanghong Shen 等WWW 2023 · 被引用 103 次
- Deep Attentive Model for Knowledge TracingXinping Wang, Liangyu Chen, Min ZhangAAAI 2023 · 被引用 10 次
- Cognitive Fluctuations Enhanced Attention Network for Knowledge TracingMingliang Hou, Xueyi Li, Teng Guo, Zitao Liu 等AAAI 2025 · 被引用 12 次
- HD-KT: Advancing Robust Knowledge Tracing via Anomalous Learning Interaction DetectionHaiping Ma, Yong Yang, Chuan Qin, Xiaoshan Yu 等WWW 2024 · 被引用 32 次
