Continual Optimization with Symmetry Teleportation for Multi-Task Learning
Zhipeng Zhou, Ziqiao Meng, Pengcheng Wu, Peilin Zhao, Chunyan Miao
摘要
Multi-task learning (MTL) is a widely explored paradigm that enables the simultaneous learning of multiple tasks using a single model. Despite numerous solutions, the key issues of optimization conflict and task imbalance remain under-addressed, limiting performance. Unlike existing optimization-based approaches that typically reweight task losses or gradients to mitigate conflicts or promote progress, we propose a novel approach based on Continual Optimization with Symmetry Teleportation (COST). During MTL optimization, when an optimization conflict arises, we seek an alternative loss-equivalent point on the loss landscape to reduce conflict. Specifically, we utilize a low-rank adapter (LoRA) to facilitate this practical teleportation by designing convergent, loss-invariant objectives. Additionally, we introduce a historical trajectory reuse strategy to continually leverage the benefits of advanced optimizers. Extensive experiments on multiple mainstream datasets demonstrate the effectiveness of our approach. COST is a plug-and-play solution that enhances a wide range of existing MTL methods. When integrated with state-of-the-art methods, COST achieves superior performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Exploring Tradeoffs through Mode Connectivity for Multi-Task LearningZhipeng Zhou, Ziqiao Meng, Pengcheng Wu, Peilin Zhao 等NeurIPS 2025 · 被引用 2 次
- From Gradient Volume to Shapley Fairness: Towards Fair Multi-Task LearningXiao Wang, Yuying Han, Dazi Li, Fei Zhang 等ICLR 2026
它引用的顶会 Paper14
- QLoRA: Efficient Finetuning of Quantized LLMsTim Dettmers, Artidoro Pagnoni, Ari Holtzman, Luke ZettlemoyerNeurIPS 2023 · 被引用 5,863 次
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 被引用 1,861 次
- Conflict-Averse Gradient Descent for Multi-task learningBo Liu, Xingchao Liu, Xiaojie Jin, Peter Stone 等NeurIPS 2021 · 被引用 686 次
- Just Pick a Sign: Optimizing Deep Multitask Models with Gradient Sign DropoutZhao Chen, Jiquan Ngiam, Yanping Huang, Thang Luong 等NeurIPS 2020 · 被引用 313 次
相关 Paper
- Conflict-Buffering Optimization by Symmetry Teleportation for Deep Long-Tailed RecognitionMianzimei Yang, Zhipeng Zhou, Jin Zhang, Yuanhao Pu 等ACM MM 2025
- Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous LearnerHanwen Zhong, Jiaxin Chen, Yutong Zhang, Di Huang 等NeurIPS 2024 · 被引用 9 次
- Scalable Multi-Task Low-Rank Model AdaptationZichen Tian, Antoine Ledent, Qianru SunICLR 2026
- ICM-Fusion: In-Context Meta-Optimized LoRA Fusion for Multi-Task AdaptationYihua Shao, Xiaofeng Lin, Xinwei Long, Siyu Chen 等AAAI 2026 · 被引用 8 次
- SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space SplittingHaomiao Qiu, Miao Zhang, Ziyue Qiao, Weili Guan 等ICLR 2026 · 被引用 10 次
