Bone Soups: A Seek-and-Soup Model Merging Approach for Controllable Multi-Objective Generation
Guofu Xie, Xiao Zhang, Ting Yao, Yunsheng Shi
摘要
User information needs are often highly diverse and varied. A key challenge in current research is how to achieve controllable multi-objective generation while enabling rapid adaptation to accommodate diverse user demands during test time. Existing solutions, such as Rewarded Soup, focus on merging language models individually tuned on single objectives. While easy to implement and widely used, these approaches face limitations in achieving optimal performance due to their disregard for the impacts of competing objectives on model tuning. To address this issue, we propose Bone Soup, a novel model merging approach that first seeks a series of backbone models by considering the impacts of multiple objectives and then makes the soup (i.e., merge the backbone models). Specifically, Bone Soup begins by training multiple backbone models for different objectives using multi-objective reinforcement learning. Each backbone model is guided by a combination of backbone reward signals. To ensure that these models are optimal for the Pareto front, the backbone rewards are crafted by combining standard reward functions into basis vectors, which can then be modified through a rulebased construction method. Bone Soup leverages a symmetric circulant matrix mapping to generate the merging coefficients, which are used to merge the backbone models according to user preferences. Extensive experimental results demonstrate that Bone Soup exhibits strong controllability and Pareto optimality in controllable multi-objective generation, providing a more effective and efficient approach to addressing diverse user needs at test time. Code is available at https://github. com/andyclsr/BoneSoups . Restructuring Backbone Models Testing Stage 𝜆 1 𝜆 2 𝜆 3 Testing Stage 0.5 0.3 0.2 SFT Model 𝒉𝟏 = 𝜷 * 𝑹𝟏 + 𝟏 -𝜷 𝟐 * 𝑹𝟐 + 𝟏 -𝜷 𝟐 * 𝑹𝟑 𝒉𝟐 = 𝟏 -𝜷 𝟐 * 𝑹𝟏 + 𝜷 * 𝑹𝟐 + 𝟏 -𝜷 𝟐 * 𝑹𝟑 𝒉𝟑 = 𝟏 -𝜷 𝟐 * 𝑹𝟏 + 𝟏 -𝜷 𝟐 * 𝑹𝟐 + 𝜷 * 𝑹𝟑
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Steerable Adversarial Scenario Generation through Test-Time Preference AlignmentTong Nie, Yuewen Mei, Yihong Tang, Junlin He 等ICLR 2026 · 被引用 8 次
- OrthAlign: Orthogonal Subspace Decomposition for Non-Interfering Multi-Objective AlignmentLiang Lin, Zhihao Xu, Junhao Dong, Jian Zhao 等ICLR 2026 · 被引用 7 次
- Multi-Value Alignment for LLMs via Value Decorrelation and ExtrapolationHefei Xu, Le Wu, Chen Cheng, Hao LiuAAAI 2026 · 被引用 4 次
- MapReduce LoRA: Advancing the Pareto Front in Multi-Preference Optimization for Generative ModelsChieh-Yun Chen, Zhonghao Wang, Qi Chen, Zhifan Ye 等CVPR 2026 · 被引用 3 次
- Self-Guided Alignment: Adaptive Preference Sensing for Multi-Objective GenerationNing Wang, Zhanyang Liu, Taotao Zhou, Xinrui Zhang 等ACL 2026
它引用的顶会 Paper21
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs 等ICML 2022 · 被引用 1,464 次
- TIES-Merging: Resolving Interference When Merging ModelsPrateek Yadav, Derek Tam, Leshem Choshen, Colin A. Raffel 等NeurIPS 2023 · 被引用 999 次
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 被引用 750 次
- Safe RLHF: Safe Reinforcement Learning from Human FeedbackJosef Dai, Xuehai Pan, Ruiyang Sun, Jiaming Ji 等ICLR 2024 · 被引用 656 次
相关 Paper
- Rewarded soups: towards Pareto-optimal alignment by interpolating weights fine-tuned on diverse rewardsAlexandre Ramé, Guillaume Couairon, Corentin Dancette, Jean-Baptiste Gaya 等NeurIPS 2023 · 被引用 295 次
- Objective Soups: Multilingual Multi-Task Modeling for Speech ProcessingA. F. M. Saif, Lisha Chen, Xiaodong Cui, Songtao Lu 等NeurIPS 2025
- PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward ModelBaijiong Lin, Weisen Jiang, Yuancheng Xu, Hao Chen 等ICML 2025
- HM3: Hierarchical Multi-Objective Model Merging for Pretrained ModelsYu Zhou, Xingyu Wu, Jibin Wu, Liang Feng 等NeurIPS 2025 · 被引用 14 次
- From Parameter to Representation: A Closed-Form Approach for Controllable Model MergingJialin Wu, Jian Yang, Handing Wang, Jiajun Wen 等AAAI 2026
