SCE: Scalable Network Embedding from Sparsest Cut
Shengzhong Zhang, Zengfeng Huang, Haicang Zhou, Ziang Zhou
Abstract
Large-scale network embedding is to learn a latent representation for each node in an unsupervised manner, which captures inherent properties and structural information of the underlying graph. In this field, many popular approaches are influenced by the skip-gram model from natural language processing. Most of them use a contrastive objective to train an encoder which forces the embeddings of similar pairs to be close and embeddings of negative samples to be far. A key of success to such contrastive learning methods is how to draw positive and negative samples. While negative samples that are generated by straightforward random sampling are often satisfying, methods for drawing positive examples remains a hot topic. In this paper, we propose SCE for unsupervised network embedding only using negative samples for training. Our method is based on a new contrastive objective inspired by the well-known sparsest cut problem. To solve the underlying optimization problem, we introduce a Laplacian smoothing trick, which uses graph convolutional operators as low-pass filters for smoothing node representations. The resulting model consists of a GCN-type structure as the encoder and a simple loss function. Notably, our model does not use positive samples but only negative samples for training, which not only makes the implementation and tuning much easier, but also reduces the training time significantly. Finally, extensive experimental studies on real world data sets are conducted. The results clearly demonstrate the advantages of our new model in both accuracy and scalability compared to strong baselines such as GraphSAGE, G2G and DGI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b36d3eba-7714-4388-81b0-4650e38c0351Cited by top-tier papers6
- Learning Based Proximity Matrix Factorization for Node EmbeddingXingyi Zhang, Kun Xie, Sibo Wang, Zengfeng HuangKDD 2021 · 30 citations
- Nonlinear Feature Diffusion on HypergraphsKonstantin Prokopchik, Austin R. Benson, Francesco TudiscoICML 2022 · 25 citations
- Road Network Representation Learning with the Third Law of GeographyHaicang Zhou, Weiming Huang, Yile Chen, Tiantian He et al.NeurIPS 2024 · 23 citations
- StructComp: Substituting propagation with Structural Compression in Training Graph Contrastive LearningShengzhong Zhang, Wenjie Yang, Xinyuan Cao, Hongwei Zhang et al.ICLR 2024 · 6 citations
- PSMC: Provable and Scalable Algorithms for Motif Conductance Based Graph ClusteringLonglong Lin, Tao Jia, Zeli Wang, Jin Zhao et al.KDD 2024 · 3 citations
Builds on1
Related papers
- Self-Contrastive Graph Diffusion NetworkYixuan Ma, Kun ZhanACM MM 2023 · 15 citations
- Adaptive Graph Encoder for Attributed Graph EmbeddingGanqu Cui, Jie Zhou, Cheng Yang, Zhiyuan LiuKDD 2020 · 224 citations
- Self-Supervised Representation Learning via Latent Graph PredictionYaochen Xie, Zhao Xu, Shuiwang JiICML 2022 · 43 citations
- Does GCL Need a Large Number of Negative Samples? Enhancing Graph Contrastive Learning with Effective and Efficient Negative SamplingYongqi Huang, Jitao Zhao, Dongxiao He, Di Jin et al.AAAI 2025 · 11 citations
- Revisiting Positive Samples in Graph Contrastive Learning: From the Perspective of Message PassingLianze Shan, Ningchong Wang, Jitao Zhao, Di Jin et al.ICML 2026
