FedMEKT: Distillation-based embedding knowledge transfer for multimodal federated learning

Le, Q. Huy; Nguyen, Huu Nhat Minh; Thwal, Chu Myaet; Qiao, Yu; Zhang, Chaoning; Hong, Choong Seon

Please use this identifier to cite or link to this item: https://elib.vku.udn.vn/handle/123456789/5802

Full metadata record

DC Field	Value	Language
dc.contributor.author	Le, Q. Huy	-
dc.contributor.author	Nguyen, Huu Nhat Minh	-
dc.contributor.author	Thwal, Chu Myaet	-
dc.contributor.author	Qiao, Yu	-
dc.contributor.author	Zhang, Chaoning	-
dc.contributor.author	Hong, Choong Seon	-
dc.date.accessioned	2025-11-13T09:54:14Z	-
dc.date.available	2025-11-13T09:54:14Z	-
dc.date.issued	2025-03	-
dc.identifier.uri	https://doi.org/10.1016/j.neunet.2024.107017	-
dc.identifier.uri	https://elib.vku.udn.vn/handle/123456789/5802	-
dc.description	Neural Networks; Volume 183, 107017	vi_VN
dc.description.abstract	Federated learning (FL) enables a decentralized machine learning paradigm for multiple clients to collaboratively train a generalized global model without sharing their private data. Most existing works have focused on designing FL systems for unimodal data, limiting their potential to exploit valuable multimodal data for future personalized applications. Moreover, the majority of FL approaches still rely on labeled data at the client side, which is often constrained by the inability of users to self-annotate their data in real-world applications. In light of these limitations, we propose a novel multimodal FL framework, namely FedMEKT, based on a semi-supervised learning approach to leverage representations from different modalities. To address the challenges of modality discrepancy and labeled data constraints in existing FL systems, our proposed FedMEKT framework comprises local multimodal autoencoder learning, generalized multimodal autoencoder construction, and generalized classifier learning. Bringing this concept into the proposed framework, we develop a distillation-based multimodal embedding knowledge transfer mechanism which allows the server and clients to exchange joint multimodal embedding knowledge extracted from a multimodal proxy dataset. Specifically, our FedMEKT iteratively updates the generalized global encoders with joint multimodal embedding knowledge from participating clients through upstream and downstream multimodal embedding knowledge transfer for local learning. Through extensive experiments on four multimodal datasets, we demonstrate that FedMEKT not only achieves superior global encoder performance in linear evaluation but also guarantees user privacy for personal data and model parameters while demanding less communication cost than other baselines.	vi_VN
dc.language.iso	en	vi_VN
dc.publisher	Elsevier	vi_VN
dc.subject	Federated learning	vi_VN
dc.subject	personal data	vi_VN
dc.subject	decentralized machine learning	vi_VN
dc.subject	FL approaches	vi_VN
dc.subject	autoencoder	vi_VN
dc.title	FedMEKT: Distillation-based embedding knowledge transfer for multimodal federated learning	vi_VN
dc.type	Working Paper	vi_VN
Appears in Collections:	NĂM 2025

Files in This Item:

Sign in to read

Show simple item record