PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 26, 2026IEEE Transactions on Neural Networks and Learning Systems3 citations

Incomplete Multimodal Federated Learning via Masking and Contrasting Prototypes

View Full Paper
GBGuangyin BaoQZQiang ZhangDMDuoqian Miao

Key Points

  • The aim is to address performance issues in multimodal federated learning caused by missing modality information.
  • Developed a novel framework to handle task drift and performance degradation due to modality missingness.
  • Constructed a prototype library to use as masks for missing modalities during training and inference.
  • Created a task-calibrated training loss and a model-agnostic inference strategy.
  • Implemented a proximal term using prototype contrastive learning for better integration of global information.
  • Demonstrated improved inference performance across various modality missingness settings.
  • Achieved a 23.8% increase in performance during modality-incomplete inference compared to existing methods.

Abstract

In real-world scenarios, random modality missingness in multimodal federated learning (mFL) poses a significant challenge, diminishing the performance of global model inference. However, existing mFL methods are predominantly limited to simple scenarios that typically involve participant clients restricted to either a single modality or multimodal clients with complete modalities. They employ modality-specific encoders on each client and train modality fusion modules on the server, leading to severe task drift between clients and server, and struggling to generalize effectively in intricate modality-missing scenarios. To this end, we present a novel mFL framework to alleviate the task drift and performance degradation resulting from modality missingness during both training and inference. Inspired by prototype learning using the highly generalized proxy of specific information, we elaborately construct a prototype library to enhance FedAvg-based federated learning (FL). Naturally, we utilize prototypes as masks representing missing modalities to compensate for the missingness of modality information, formulating a task-calibrated training loss and devising a model-agnostic modality-incomplete inference strategy. In addition, a proximal term based on prototype contrastive learning is constructed to integrate interclient global information into each client, therefore enhancing local training. We conduct extensive experiments to evaluate our mFL framework, demonstrating its state-of-the-art performance across a series of missingness settings. Specifically, compared with existing mFL methods, our mFL framework improves inference performance under different modality missingness rates during training and by 23.8% during modality-incomplete inference.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Bao et al. (2026) studied this question.

synapsesocial.com/papers/69c4cc37fdc3bde4489176c6https://doi.org/10.1109/tnnls.2026.3658522
Ask AI
Helpful
Bookmark
Share
View Full Paper