IterIS: Iterative Inference-Solving Alignment for LoRA Merging
Hongxu Chen, Zhen Wang, Runshi Li, Bowei Zhu, Long Chen
Abstract
a [V 2 ] cat standing by a [V 1 ] barn LoRA barn LoRA cat Adapter cat+barn Adapter NEG+POS (b) Multi-Style Caption LoRA NEG LoRA POS (c) Multiple NLP Tasks Integration LoRA A <Human>: Question Type A Adapter A+B Question Type A Question Type B Answer Answer ✅ Caption (NEG): a dead man sitting on a couch with a laptop and a dog + Caption (POS): a pretty woman in a red jacket skiing down a snowy hill LoRA B <Human>: Question Type B <Robot> : Answer Type B ✅ <Robot> : Answer Type A LLM Caption (NEG): a group of stupid people playing baseball on a field Caption (POS): a good team of baseball players standing around home plate during a game Figure 1. Overview of the application of our method (IterIS) across multiple domains. Our general method is adaptable for merging LoRAs in various contexts. IterIS can be applied to (a) text-to-image diffusion models for multi-concept customization, (b) vision-language models for multi-style caption generation, and (c) large language models for multiple NLP tasks integration.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- DiffGraph: An Automated Agent-driven Model Merging Framework for In-the-Wild Text-to-Image GenerationZhuoling Li, Hossein Rahmani, Jiarui Zhang, Yu Xue et al.CVPR 2026 · 5 citations
- SSR-Merge: Subspace Signal Routing for Training-Free LoRA Merging in Diffusion ModelsZhengxuan Wei, Yi Dong, Zonghui Li, Xianhui Lin et al.ICML 2026
- Compress then Merge: From Multiple LoRAs into One Low-Rank AdapterZhengbao He, Ruiqi Ding, Zhehao Huang, Ruikai Yang et al.ICML 2026
Builds on12
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 6,549 citations
- TIES-Merging: Resolving Interference When Merging ModelsPrateek Yadav, Derek Tam, Leshem Choshen, Colin A. Raffel et al.NeurIPS 2023 · 999 citations
- Mix-of-Show: Decentralized Low-Rank Adaptation for Multi-Concept Customization of Diffusion ModelsYuchao Gu, Xintao Wang, Jay Zhangjie Wu, Yujun Shi et al.NeurIPS 2023 · 333 citations
Related papers
- More Than Catastrophic Forgetting: Integrating General Capabilities For Domain-Specific LLMsChengyuan Liu, Yangyang Kang, Shihang Wang, Lizhi Qing et al.EMNLP 2024 · 6 citations
- Label-Free Cross-Task LoRA Merging with Null-Space CompressionWonyoung Lee, Wooseong Jeong, Kuk-Jin YoonCVPR 2026 · 3 citations
- Training-free LLM Merging for Multi-task LearningZichuan Fu, Xian Wu, Yejing Wang, Wanyu Wang et al.ACL 2025
- LoRACLR: Contrastive Adaptation for Customization of Diffusion ModelsEnis Simsar, Thomas Hofmann, Federico Tombari, Pinar YanardagCVPR 2025
- Empower Vision Applications with LoRA LMMLiang Mi, Weijun Wang, Wenming Tu, Qingfeng He et al.EuroSys 2025 · 2 citations
