Cross-Modal Federated Human Activity Recognition via Modality-Agnostic and Modality-Specific Representation Learning
Xiaoshan Yang, Baochen Xiong, Yi Huang, Changsheng Xu
Abstract
In this paper, we propose a new task of cross-modal federated human activity recognition (CMF-HAR), which is conducive to promote the large-scale use of the HAR model on more local devices. To address the new task, we propose a feature-disentangled activity recognition network (FDARN), which has five important modules of altruistic encoder, egocentric encoder, shared activity classifier, private activity classifier and modality discriminator. The altruistic encoder aims to collaboratively embed local instances on different clients into a modality-agnostic feature subspace. The egocentric encoder aims to produce modality-specific features that cannot be shared across clients with different modalities. The modality discriminator is used to adversarially guide the parameter learning of the altruistic and egocentric encoders. Through decentralized optimization with a spherical modality discriminative loss, our model can not only generalize well across different clients by leveraging the modality-agnostic features but also capture the modality-specific discriminative characteristics of each client. Extensive experiment results on four datasets demonstrate the effectiveness of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1f227bb7-b4d2-4fd4-8930-c5afa5f2691aCited by top-tier papers5
- Learning Unseen Modality InteractionYunhua Zhang, Hazel Doughty, Cees SnoekNeurIPS 2023 · 16 citations
- Enhancing Storage and Computational Efficiency in Federated Multimodal Learning for Large-Scale ModelsZixin Zhang, Fan Qi, Changsheng XuICML 2024 · 5 citations
- X-FLoRA: Cross-modal Federated Learning with Modality-expert LoRA for Medical VQAMin Hyuk Kim, Changheon Kim, Seok Bong YooEMNLP 2025 · 2 citations
- Sharp Eyes and Memory for VideoLLMs: Information-Aware Visual Token Pruning for Efficient and Reliable VideoLLM ReasoningJialong Qin, Xin Zou, Di Lu, Yibo Yan et al.AAAI 2026
- Adaptive Hyper-graph Aggregation for Modality-Agnostic Federated LearningQ. Fan, L. ShuaiCVPR 2024
Builds on8
- Personalized Federated Learning with Moreau EnvelopesCanh T. Dinh, Nguyen Hoang Tran, Tuan Dung NguyenNeurIPS 2020 · 1,542 citations
- Federated Learning with Matched AveragingHongyi Wang, Mikhail Yurochkin, Yuekai Sun, Dimitris S. Papailiopoulos et al.ICLR 2020 · 1,368 citations
- Personalized Federated Learning with Theoretical Guarantees: A Model-Agnostic Meta-Learning ApproachAlireza Fallah, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2020 · 1,354 citations
- FedBN: Federated Learning on Non-IID Features via Local Batch NormalizationXiaoxiao Li, Meirui Jiang, Xiaofei Zhang, Michael Kamp et al.ICLR 2021 · 1,166 citations
- Exploiting Shared Representations for Personalized Federated LearningLiam Collins, Hamed Hassani, Aryan Mokhtari, Sanjay ShakkottaiICML 2021 · 1,081 citations
Related papers
- Meta-HAR: Federated Representation Learning for Human Activity RecognitionChenglin Li, Di Niu, Bei Jiang, Xiao Zuo et al.WWW 2021 · 129 citations
- MFC: Mixed Federated Clustering based on Cross-modal Feature DecouplingXiaxia He, Boyue Wang, Junbin Gao, Yongli Hu et al.KDD 2026
- Adversarial Multi-view Networks for Activity RecognitionLei Bai, Lina Yao, Xianzhi Wang, Salil S. Kanhere et al.UbiComp 2020 · 41 citations
- HMGAN: A Hierarchical Multi-Modal Generative Adversarial Network Model for Wearable Human Activity RecognitionLing Chen, Rong Hu, Menghan Wu, Xin ZhouUbiComp 2023 · 23 citations
- Disentangled Representation Learning for Multimodal Emotion RecognitionDingkang Yang, Shuai Huang, Haopeng Kuang, Yangtao Du et al.ACM MM 2022 · 260 citations
