Micro-video Tagging via Jointly Modeling Social Influence and Tag Relation
Xiao Wang, Tian Gan, Yinwei Wei, Jianlong Wu, Dai Meng, Liqiang Nie
Abstract
The last decade has witnessed the proliferation of micro-videos on various user-generated content platforms. According to our statistics, around 85.7% of micro-videos lack annotation. In this paper, we focus on annotating micro-videos with tags. Existing methods mostly focus on analyzing video content, neglecting users' social influence and tag relation. Meanwhile, existing tag relation construction methods suffer from either deficient performance or low tag coverage. To jointly model social influence and tag relation, we formulate micro-video tagging as a link prediction problem in a constructed heterogeneous network. Specifically, the tag relation (represented by tag ontology) is constructed in a semi-supervised manner. Then, we combine tag relation, video-tag annotation, and user follow relation to build the network. Afterward, a better video and tag representation are derived through Behavior Spread modeling and visual and linguistic knowledge aggregation. Finally, the semantic similarity between each micro-video and all candidate tags is calculated in this video-tag network. Extensive experiments on industrial datasets of three verticals verify the superiority of our model compared with several state-of-the-art baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- RTQ: Rethinking Video-language Understanding Based on Image-text ModelXiao Wang, Yaoyu Li, Tian Gan, Zheng Zhang et al.ACM MM 2023 · 14 citations
- Temporal Sentence Grounding in Streaming VideosTian Gan, Xiao Wang, Yan Sun, Jianlong Wu et al.ACM MM 2023 · 5 citations
- An Inverse Partial Optimal Transport Framework for Music-guided Trailer GenerationYutong Wang, Sidan Zhu, Hongteng Xu, Dixin LuoACM MM 2024 · 2 citations
Builds on8
- Graph-Refined Convolutional Network for Multimedia Recommendation with Implicit FeedbackYinwei Wei, Xiang Wang, Liqiang Nie, Xiangnan He et al.ACM MM 2020 · 374 citations
- Are we really making much progress?: Revisiting, benchmarking and refining heterogeneous graph neural networksQingsong Lv, Ming Ding, Qiang Liu, Yuxiang Chen et al.KDD 2021 · 249 citations
- Cross-Modality Attention with Semantic Graph Embedding for Multi-Label ClassificationRenchun You, Zhiyao Guo, Lei Cui, Xiang Long et al.AAAI 2020 · 221 citations
- Multi-Label Classification with Label Graph SuperimposingYa Wang, Dongliang He, Fu Li, Xiang Long et al.AAAI 2020 · 192 citations
- Adversarial Multimodal Representation Learning for Click-Through Rate PredictionXiang Li, Chao Wang, Jiwei Tan, Xiaoyi Zeng et al.WWW 2020 · 61 citations
Related papers
- Zero-shot Recommendation: Towards Class Semantic Relation Learning for Inferring Labels of Unseen Micro-videosJunyang Chen, Huan Wang, Yirui Wu, Qiuzhen Lin et al.AAAI 2026
- Adaptive Anti-Bottleneck Multi-Modal Graph Learning Network for Personalized Micro-video RecommendationDesheng Cai, Shengsheng Qian, Quan Fang, Jun Hu et al.ACM MM 2022 · 19 citations
- Learning Fine-grained User Interests for Micro-video RecommendationYu Shang, Chen Gao, Jiansheng Chen, Depeng Jin et al.SIGIR 2023 · 19 citations
- Weakly Supervised Attention for Hashtag Recommendation using Graph DataAmin Javari, Zhankui He, Zijie Huang, Jeetu Raj et al.WWW 2020 · 22 citations
- Learning User Representations for Open Vocabulary Image Hashtag PredictionThibaut DurandCVPR 2020
