Making Logic a First-Class Citizen in Generative ML for Networking
Hongyu Hè, Minhao Jin, Maria Apostolaki
摘要
Generative ML models are increasingly popular in networking for tasks such as telemetry imputation, prediction, and synthetic trace generation. Despite their capabilities, they suffer from two shortcomings: (i) their output is often visibly violating well-known networking rules, which undermines their trustworthiness; and (ii) they are difficult to control, frequently requiring retraining even for minor changes. To address these limitations and unlock the benefits of generative models for networking, we propose a new paradigm for integrating explicit network knowledge, in the form of first-order logic rules, into ML models used for networking tasks. Rules capture well-known relationships among observed signals, e.g., that increased latency precedes packet loss. While the idea is conceptually straightforward, its realization is challenging: networking knowledge is rarely formalized into rules, and naively injecting rules into ML models often hampers their effectiveness. This paper introduces NetNomos, a multi-stage framework that (i) learns rules directly from data (e.g., measurements); (ii) filters them to select semantically meaningful ones; and (iii) enforces them through collaborative generation between an ML model and a Satisfiability Modulo Theories (SMT) solver. %We evaluate NetNomos both component-wise and end-to-end across four diverse network datasets. We show that NetNomos learns diverse, meaningful rules from four real-world datasets and is 1.6--6.5 more scalable than DuoAI, a state-of-the-art (SOTA) rule-learning method. By enforcing these rules on a generic GPT-2 model, NetNomos achieves performance on par with or even surpassing specialized SOTA systems such as Zoom2Net and NetShare across three networking tasks: telemetry imputation, traffic forecasting, and synthetic data generation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper21
- Synchromesh: Reliable Code Generation from Pre-trained Language ModelsGabriel Poesia, Alex Polozov, Vu Le, Ashish Tiwari 等ICLR 2022 · 被引用 200 次
- Charformer: Fast Character Transformers via Gradient-based Subword TokenizationYi Tay, Vinh Q. Tran, Sebastian Ruder, Jai Prakash Gupta 等ICLR 2022 · 被引用 198 次
- Practical GAN-based synthetic IP header trace generation using NetShareYucheng Yin, Zinan Lin, Minhao Jin, Giulia Fanti 等SIGCOMM 2022 · 被引用 106 次
- Validating SMT solvers via semantic fusionDominik Winterer, Chengyu Zhang, Zhendong SuPLDI 2020 · 被引用 80 次
- AI/ML for Network Security: The Emperor has no ClothesArthur Selle Jacobs, Roman Beltiukov, Walter Willinger, Ronaldo A. Ferreira 等CCS 2022 · 被引用 76 次
相关 Paper
- NetLLM: Adapting Large Language Models for NetworkingDuo Wu, Xianda Wang, Yaqi Qiao, Zhi Wang 等SIGCOMM 2024 · 被引用 162 次
- Zoom2Net: Constrained Network Telemetry ImputationFengchen Gong, Divya Raghunathan, Aarti Gupta, Maria ApostolakiSIGCOMM 2024 · 被引用 11 次
- ARI-LLM: Autoregressive Imputation for Network Traffic Matrix via Large Language ModelsFenglin Yan, Kaiwen Jiang, Yan Qiao, Meng Li 等INFOCOM 2026 · 被引用 1 次
- Net-Ev2: A Generative Simulator for Network Event EvolutionGuangyu Wang, Zhaonan WangKDD 2026 · 被引用 1 次
- Neural Semantic Parsing in Low-Resource Settings with Back-Translation and Meta-LearningYibo Sun, Duyu Tang, Nan Duan, Yeyun Gong 等AAAI 2020 · 被引用 25 次
