Survey of Multimodal Federated Learning: Exploring Data Integration, Challenges, and Future Directions

Mumin Adam, Abdullatif Albaseer, Uthman Baroudi, Mohamed Abdallah*

*Corresponding author for this work

Research output: Contribution to journalArticlepeer-review

Abstract

The rapidly expanding demand for intelligent wireless applications and the Internet of Things (IoT) requires advanced system designs to handle multimodal data effectively while ensuring user privacy and data security. Traditional machine learning (ML) models rely on centralized architectures, which, while powerful, often present significant privacy risks due to the centralization of sensitive data. Federated Learning (FL) is a promising decentralized alternative for addressing these issues. However, FL predominantly handles unimodal data, which limits its applicability in environments where devices collect and process various data types such as text, images, and sensor output. To address this limitation, Multimodal FL (MMFL) integrates multiple data modalities, enabling a richer and more holistic understanding of data. In this survey, we explore the challenges and advancements in MMFL, including data representation, fusion techniques, and cross-modal learning strategies. We present a comprehensive taxonomy of MMFL, outlining critical challenges such as modality imbalance, fusion complexity, and security concerns. Additionally, we highlight the role of transformers in MMFL by leveraging their powerful attention mechanisms to process multimodal data in a federated setting. Finally, we discuss various applications of MMFL, including healthcare, human activity recognition, and emotion recognition, and propose future research directions for improving the scalability and robustness of MMFL systems in real-world scenarios.

Original languageEnglish
Pages (from-to)2510-2538
Number of pages29
JournalIEEE Open Journal of the Communications Society
Volume6
DOIs
Publication statusPublished - 2025

Keywords

  • Accuracy
  • Computational modeling
  • Cross-modal
  • Data fusion
  • Data models
  • Data privacy
  • Distributed databases
  • Federated learning
  • Internet of Things
  • Multimodal FL
  • Multimodal federated transformer learning
  • Scalability
  • Surveys
  • Transformers
  • multimodal FL communication intelligent IoT applications

Fingerprint

Dive into the research topics of 'Survey of Multimodal Federated Learning: Exploring Data Integration, Challenges, and Future Directions'. Together they form a unique fingerprint.

Cite this