Lune

ICCV2025Top-tier venue

G2D: Boosting Multimodal Learning with Gradient-Guided Distillation

Mohammed Rakib, Arunkumar Bagavathi

2025Year
1Citations

Abstract

Multimodal learning aims to leverage information from diverse data modalities to achieve more comprehensive performance. However, conventional multimodal models often suffer from modality imbalance, where one or a few modalities dominate model optimization, leading to suboptimal feature representation and underutilization of weak modalities. To address this challenge, we introduce GradientGuided Distillation (G2D)(\mathrm{G}^{2} \mathrm{D}), a knowledge distillation framework that optimizes the multimodal model with a custombuilt loss function that fuses both unimodal and multimodal objectives. G2DG^{2} D further incorporates a dynamic sequential modality prioritization (SMP) technique in the learning process to ensure each modality leads the learning process, avoiding the pitfall of stronger modalities overshadowing weaker ones. We validate G2DG^{2} D on multiple realworld datasets and show that G2DG^{2} D amplifies the significance of weak modalities while training and outperforms state-of-the-art methods in classification and regression tasks. Our code is available here.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext ff311c9e-94ed-4076-9174-6b260d69c882

Builds on22

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines