Federated Learning on Heterogenous Data using Chest CT

03/23/2023
by   Edward H. Lee, et al.
0

Large data have accelerated advances in AI. While it is well known that population differences from genetics, sex, race, diet, and various environmental factors contribute significantly to disease, AI studies in medicine have largely focused on locoregional patient cohorts with less diverse data sources. Such limitation stems from barriers to large-scale data share in medicine and ethical concerns over data privacy. Federated learning (FL) is one potential pathway for AI development that enables learning across hospitals without data share. In this study, we show the results of various FL strategies on one of the largest and most diverse COVID-19 chest CT datasets: 21 participating hospitals across five continents that comprise >10,000 patients with >1 million images. We present three techniques: Fed Averaging (FedAvg), Incremental Institutional Learning (IIL), and Cyclical Incremental Institutional Learning (CIIL). We also propose an FL strategy that leverages synthetically generated data to overcome class imbalances and data size disparities across centers. We show that FL can achieve comparable performance to Centralized Data Sharing (CDS) while maintaining high performance across sites with small, underrepresented data. We investigate the strengths and weaknesses for all technical approaches on this heterogeneous dataset including the robustness to non-Independent and identically distributed (non-IID) diversity of data. We also describe the sources of data heterogeneity such as age, sex, and site locations in the context of FL and show how even among the correctly labeled populations, disparities can arise due to these biases.

READ FULL TEXT

page 13

page 14

page 15

page 16

page 20

page 21

page 22

research
06/07/2023

Phoenix: A Federated Generative Diffusion Model

Generative AI has made impressive strides in enabling users to create di...
research
11/18/2021

Advancing COVID-19 Diagnosis with Privacy-Preserving Collaboration in Artificial Intelligence

Artificial intelligence (AI) provides a promising substitution for strea...
research
08/22/2023

Federated Learning on Patient Data for Privacy-Protecting Polycystic Ovary Syndrome Treatment

The field of women's endocrinology has trailed behind data-driven medica...
research
08/07/2023

The Prospect of Enhancing Large-Scale Heterogeneous Federated Learning with Transformers

Federated learning (FL) addresses data privacy concerns by enabling coll...
research
07/17/2023

Privacy-preserving patient clustering for personalized federated learning

Federated Learning (FL) is a machine learning framework that enables mul...
research
04/22/2022

Application of Federated Learning in Building a Robust COVID-19 Chest X-ray Classification Model

While developing artificial intelligence (AI)-based algorithms to solve ...
research
10/28/2022

Federated Learning for Chronic Obstructive Pulmonary Disease Classification with Partial Personalized Attention Mechanism

Chronic Obstructive Pulmonary Disease (COPD) is the fourth leading cause...

Please sign up or login with your details

Forgot password? Click here to reset