REVIEW 3 cited by
Examining Modality Incongruity in Multimodal Federated Learning for Medical Vision and Language-based Disease Detection
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Multimodal Federated Learning (MMFL) utilizes multiple modalities in each client to build a more powerful Federated Learning (FL) model than its unimodal counterpart. However, the impact of missing modality in different clients, also called modality incongruity, has been greatly overlooked. This paper, for the first time, analyses the impact of modality incongruity and reveals its connection with data heterogeneity across participating clients. We particularly inspect whether incongruent MMFL with unimodal and multimodal clients is more beneficial than unimodal FL. Furthermore, we examine three potential routes of addressing this issue. Firstly, we study the effectiveness of various self-attention mechanisms towards incongruity-agnostic information fusion in MMFL. Secondly, we introduce a modality imputation network (MIN) pre-trained in a multimodal client for modality translation in unimodal clients and investigate its potential towards mitigating the missing modality problem. Thirdly, we assess the capability of client-level and server-level regularization techniques towards mitigating modality incongruity effects. Experiments are conducted under several MMFL settings on two publicly available real-world datasets, MIMIC-CXR and Open-I, with Chest X-Ray and radiology reports.
Forward citations
Cited by 3 Pith papers
-
ProMoE-FL: Prototype-conditioned Mixture of Experts for Multimodal Federated Learning with Missing Modalities
Prototype-conditioned Mixture-of-Experts synthesizes missing modalities in federated learning and beats prior methods on heterogeneous chest X-ray clients without public data.
-
Multimodal Federated Learning With Missing Modalities through Feature Imputation Network
A federated feature imputation network that synthesizes missing modality bottleneck features improves multimodal federated learning accuracy over naive and generative baselines.
-
Robust Incomplete-Modality Alignment for Ophthalmic Disease Grading and Diagnosis via Labeled Optimal Transport
A labeled optimal transport alignment plus asymmetric fusion keeps multimodal eye disease grading accurate even when fundus or OCT is missing.
Discussion (0). Continue with ORCID to comment.