Pith. sign in

REVIEW 3 cited by

Examining Modality Incongruity in Multimodal Federated Learning for Medical Vision and Language-based Disease Detection

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.05294 v1 pith:TD7XKPZ6 submitted 2024-02-07 cs.LG cs.AIcs.CLcs.CV

classification cs.LGcs.AIcs.CLcs.CV
keywords modalityclientsincongruitymmflmultimodalunimodalfederatedlearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Multimodal Federated Learning (MMFL) utilizes multiple modalities in each client to build a more powerful Federated Learning (FL) model than its unimodal counterpart. However, the impact of missing modality in different clients, also called modality incongruity, has been greatly overlooked. This paper, for the first time, analyses the impact of modality incongruity and reveals its connection with data heterogeneity across participating clients. We particularly inspect whether incongruent MMFL with unimodal and multimodal clients is more beneficial than unimodal FL. Furthermore, we examine three potential routes of addressing this issue. Firstly, we study the effectiveness of various self-attention mechanisms towards incongruity-agnostic information fusion in MMFL. Secondly, we introduce a modality imputation network (MIN) pre-trained in a multimodal client for modality translation in unimodal clients and investigate its potential towards mitigating the missing modality problem. Thirdly, we assess the capability of client-level and server-level regularization techniques towards mitigating modality incongruity effects. Experiments are conducted under several MMFL settings on two publicly available real-world datasets, MIMIC-CXR and Open-I, with Chest X-Ray and radiology reports.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ProMoE-FL: Prototype-conditioned Mixture of Experts for Multimodal Federated Learning with Missing Modalities

    cs.CV 2026-07 conditional novelty 6.0 of 10

    Prototype-conditioned Mixture-of-Experts synthesizes missing modalities in federated learning and beats prior methods on heterogeneous chest X-ray clients without public data.

  2. Multimodal Federated Learning With Missing Modalities through Feature Imputation Network

    cs.LG 2025-05 conditional novelty 6.0 of 10

    A federated feature imputation network that synthesizes missing modality bottleneck features improves multimodal federated learning accuracy over naive and generative baselines.

  3. Robust Incomplete-Modality Alignment for Ophthalmic Disease Grading and Diagnosis via Labeled Optimal Transport

    cs.CV 2025-07 conditional novelty 5.0 of 10

    A labeled optimal transport alignment plus asymmetric fusion keeps multimodal eye disease grading accurate even when fundus or OCT is missing.

Pith tools