Bone Soups: A Seek-and-Soup Model Merging Approach for Controllable Multi-Objective Generation
Guofu Xie, Xiao Zhang, Ting Yao, Yunsheng Shi
Abstract
User information needs are often highly diverse and varied. A key challenge in current research is how to achieve controllable multi-objective generation while enabling rapid adaptation to accommodate diverse user demands during test time. Existing solutions, such as Rewarded Soup, focus on merging language models individually tuned on single objectives. While easy to implement and widely used, these approaches face limitations in achieving optimal performance due to their disregard for the impacts of competing objectives on model tuning. To address this issue, we propose Bone Soup, a novel model merging approach that first seeks a series of backbone models by considering the impacts of multiple objectives and then makes the soup (i.e., merge the backbone models). Specifically, Bone Soup begins by training multiple backbone models for different objectives using multi-objective reinforcement learning. Each backbone model is guided by a combination of backbone reward signals. To ensure that these models are optimal for the Pareto front, the backbone rewards are crafted by combining standard reward functions into basis vectors, which can then be modified through a rulebased construction method. Bone Soup leverages a symmetric circulant matrix mapping to generate the merging coefficients, which are used to merge the backbone models according to user preferences. Extensive experimental results demonstrate that Bone Soup exhibits strong controllability and Pareto optimality in controllable multi-objective generation, providing a more effective and efficient approach to addressing diverse user needs at test time. Code is available at https://github. com/andyclsr/BoneSoups . Restructuring Backbone Models Testing Stage 𝜆 1 𝜆 2 𝜆 3 Testing Stage 0.5 0.3 0.2 SFT Model 𝒉𝟏 = 𝜷 * 𝑹𝟏 + 𝟏 -𝜷 𝟐 * 𝑹𝟐 + 𝟏 -𝜷 𝟐 * 𝑹𝟑 𝒉𝟐 = 𝟏 -𝜷 𝟐 * 𝑹𝟏 + 𝜷 * 𝑹𝟐 + 𝟏 -𝜷 𝟐 * 𝑹𝟑 𝒉𝟑 = 𝟏 -𝜷 𝟐 * 𝑹𝟏 + 𝟏 -𝜷 𝟐 * 𝑹𝟐 + 𝜷 * 𝑹𝟑
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4d164e96-b453-45fa-a392-fb30112bcbe3Cited by top-tier papers6
- Steerable Adversarial Scenario Generation through Test-Time Preference AlignmentTong Nie, Yuewen Mei, Yihong Tang, Junlin He et al.ICLR 2026 · 8 citations
- OrthAlign: Orthogonal Subspace Decomposition for Non-Interfering Multi-Objective AlignmentLiang Lin, Zhihao Xu, Junhao Dong, Jian Zhao et al.ICLR 2026 · 7 citations
- Multi-Value Alignment for LLMs via Value Decorrelation and ExtrapolationHefei Xu, Le Wu, Chen Cheng, Hao LiuAAAI 2026 · 4 citations
- MapReduce LoRA: Advancing the Pareto Front in Multi-Preference Optimization for Generative ModelsChieh-Yun Chen, Zhonghao Wang, Qi Chen, Zhifan Ye et al.CVPR 2026 · 3 citations
- Self-Guided Alignment: Adaptive Preference Sensing for Multi-Objective GenerationNing Wang, Zhanyang Liu, Taotao Zhou, Xinrui Zhang et al.ACL 2026
Builds on21
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs et al.ICML 2022 · 1,464 citations
- TIES-Merging: Resolving Interference When Merging ModelsPrateek Yadav, Derek Tam, Leshem Choshen, Colin A. Raffel et al.NeurIPS 2023 · 999 citations
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 750 citations
- Safe RLHF: Safe Reinforcement Learning from Human FeedbackJosef Dai, Xuehai Pan, Ruiyang Sun, Jiaming Ji et al.ICLR 2024 · 656 citations
Related papers
- Rewarded soups: towards Pareto-optimal alignment by interpolating weights fine-tuned on diverse rewardsAlexandre Ramé, Guillaume Couairon, Corentin Dancette, Jean-Baptiste Gaya et al.NeurIPS 2023 · 295 citations
- Objective Soups: Multilingual Multi-Task Modeling for Speech ProcessingA. F. M. Saif, Lisha Chen, Xiaodong Cui, Songtao Lu et al.NeurIPS 2025
- PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward ModelBaijiong Lin, Weisen Jiang, Yuancheng Xu, Hao Chen et al.ICML 2025
- HM3: Hierarchical Multi-Objective Model Merging for Pretrained ModelsYu Zhou, Xingyu Wu, Jibin Wu, Liang Feng et al.NeurIPS 2025 · 14 citations
- From Parameter to Representation: A Closed-Form Approach for Controllable Model MergingJialin Wu, Jian Yang, Handing Wang, Jiajun Wen et al.AAAI 2026
