TrafficFormer: An Efficient Pre-trained Model for Traffic Data
Guangmeng Zhou, Xiongwen Guo, Zhuotao Liu, Tong Li, Qi Li, Ke Xu
摘要
Traffic data contains deep domain-specific knowledge, making labeling challenging, and the lack of labeled data adversely impacts the accuracy of learning-based traffic analysis. The pre-training technology is widely adopted in the fields of vision and natural language to address the problem of limited labeled data. However, the exploration in the domain of traffic analysis remains insufficient. This paper proposes an efficient pre-training model, TrafficFormer, for traffic data. In the pre-training stage, TrafficFormer introduces a fine-grained multi-classification task to enhance the representation capabilities of traffic data; in the fine-tuning stage, TrafficFormer proposes a traffic data augmentation method utilizing the random initialization feature of fields, which helps the traffic model focus on key information. We evaluate TrafficFormer using both traffic classification tasks and protocol understanding tasks. The experimental results show that TrafficFormer achieves superior performance on six traffic classification datasets, with improvements of up to 10% in the F1 score and demonstrates significantly superior protocol understanding capabilities compared to existing traffic pre-training models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- The Sweet Danger of Sugar: Debunking Representation Learning for Encrypted Traffic ClassificationYuqi Zhao, Giovanni Dettori, Matteo Boffa, Luca Vassio 等SIGCOMM 2025 · 被引用 17 次
- FENIX: Enabling In-Network DNN Inference with FPGA-Enhanced Programmable SwitchesXiangyu Gao, Tong Li, Yinchao Zhang, Ziqiang Wang 等NSDI 2026 · 被引用 12 次
- Pegasus: A Universal Framework for Scalable Deep Learning Inference on the DataplaneYinchao Zhang, Su Yao, Yong Feng, Kang Chen 等SIGCOMM 2025 · 被引用 10 次
- A Hard-Label Black-Box Evasion Attack against ML-based Malicious Traffic Detection SystemsZixuan Liu, Yi Zhao, Zhuotao Liu, Qi Li 等NDSS 2026 · 被引用 3 次
- Tracegram: Framing Trace-Level Traffic Analysis with Temporally-Aware Multiple Instance LearningJian Qu, Yuchen Zhang, Jialong Zhang, Jianfeng Li 等USENIX Security 2026
它引用的顶会 Paper31
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li 等AAAI 2020 · 被引用 4,134 次
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
相关 Paper
- Yet Another Traffic Classifier: A Masked Autoencoder Based Traffic Transformer with Multi-Level Flow RepresentationRuijie Zhao, Mingwei Zhan, Xianwen Deng, Yanhao Wang 等AAAI 2023 · 被引用 138 次
- FlowRefiner: A Robust Traffic Classification Framework against Label NoiseMingwei Zhan, Ruijie Zhao, Xianwen Deng, Zhi Xue 等NeurIPS 2025 · 被引用 2 次
- ET-BERT: A Contextualized Datagram Representation with Pre-training Transformers for Encrypted Traffic ClassificationXinjie Lin, Gang Xiong, Gaopeng Gou, Zhen Li 等WWW 2022 · 被引用 490 次
- HF-Transformer: A Non-Pretrained Encrypted Network Traffic Classification Model Based on Packet Header FieldsZhenzhen Yan, Lizhi Peng, Peiqiang Liu, Yingshuo Bao 等INFOCOM 2026
- MM4flow: A Pre-trained Multi-modal Model for Versatile Network Traffic AnalysisLuming Yang, Lin Liu, Junjie Huang, Zhuotao Liu 等CCS 2025
