FEMT-GAT: An Explainable Federated Multimodal Transformer–Graph Attention Network for Privacy-Preserving Disease Prediction

Authors

  • A. Raghavendra Rao Department of Computer Science and Engineering, Sri Chandrasekharendra Saraswathi Viswa Mahavidyalaya (SCSVMV), Enathur, Kanchipuram-631561, TamilNadu, India.
  • C. K. Gomathy Department of Computer Science and Engineering, Sri Chandrasekharendra Saraswathi Viswa Mahavidyalaya (SCSVMV), Enathur, Kanchipuram-631561, TamilNadu, India.

DOI:

https://doi.org/10.70917/ijcisim-2026-3898

Keywords:

Federated learning, multimodal learning, transformer, graph attention network, explainable artificial intelligence, differential privacy, secure aggregation, disease prediction

Abstract

The growing use of multimodal clinical data creates an opportunity to improve disease prediction, but conventional centralized learning requires hospitals to pool sensitive patient records and often provides limited interpretability. This paper presents FEMT-GAT, an explainable Federated Multimodal Transformer–Graph Attention Network for privacy-preserving disease prediction across geographically distributed healthcare institutions. The framework uses modality-specific encoders for electronic health records, laboratory measurements, demographic variables, and medical images. A cross-modal transformer learns context-dependent interactions among available modalities, while a patient similarity graph and multi-head graph attention network incorporate relational evidence from clinically comparable cases. Collaborative training is performed through federated optimization so that raw data remain at the originating institutions. Update clipping, differential privacy, encryption, and secure aggregation are incorporated to reduce information leakage. The inference module produces disease probabilities together with modality-level, feature-level, image-level, and graph-level explanations using attention rollout, modality ablation, SHAP or integrated gradients, Grad-CAM, and graph attention coefficients. The proposed architecture supports incomplete modalities through masking and modality dropout and can be adapted to binary diagnosis, multiclass classification, risk regression, and survival analysis. FEMT-GAT therefore provides a unified methodology for accurate, transparent, and privacy-aware collaborative clinical intelligence.

Downloads

Download data is not yet available.

Downloads

Published

2026-07-29

How to Cite

A. Raghavendra Rao, & C. K. Gomathy. (2026). FEMT-GAT: An Explainable Federated Multimodal Transformer–Graph Attention Network for Privacy-Preserving Disease Prediction. International Journal of Computer Information Systems and Industrial Management Applications, 18(12s), 330–353. https://doi.org/10.70917/ijcisim-2026-3898

Issue

Section

Original Articles