Distributed deep learning networks among institutions for medical imaging

Ken Chang; Niranjan Balachandar; Carson Lam; Darvin Yi; James Brown; Andrew Beers; Bruce Rosen; Daniel L Rubin; Jayashree Kalpathy-Cramer

doi:10.1093/jamia/ocy017

Distributed deep learning networks among institutions for medical imaging

J Am Med Inform Assoc. 2018 Aug 1;25(8):945-954. doi: 10.1093/jamia/ocy017.

Authors

Ken Chang¹, Niranjan Balachandar², Carson Lam², Darvin Yi², James Brown¹, Andrew Beers¹, Bruce Rosen¹, Daniel L Rubin², Jayashree Kalpathy-Cramer^{1

3}

Affiliations

¹ Athinoula A. Martinos Center for Biomedical Imaging, Department of Radiology, Massachusetts General Hospital, Charlestown, MA, 02129, USA.
² Department of Radiology and Biomedical Data Science, Stanford University, Palo Alto, CA, 94305, USA.
³ MGH and BWH Center for Clinical Data Science, Massachusetts General Hospital, Boston, MA, 02114, USA.

Abstract

Objective: Deep learning has become a promising approach for automated support for clinical diagnosis. When medical data samples are limited, collaboration among multiple institutions is necessary to achieve high algorithm performance. However, sharing patient data often has limitations due to technical, legal, or ethical concerns. In this study, we propose methods of distributing deep learning models as an attractive alternative to sharing patient data.

Methods: We simulate the distribution of deep learning models across 4 institutions using various training heuristics and compare the results with a deep learning model trained on centrally hosted patient data. The training heuristics investigated include ensembling single institution models, single weight transfer, and cyclical weight transfer. We evaluated these approaches for image classification in 3 independent image collections (retinal fundus photos, mammography, and ImageNet).

Results: We find that cyclical weight transfer resulted in a performance that was comparable to that of centrally hosted patient data. We also found that there is an improvement in the performance of cyclical weight transfer heuristic with a high frequency of weight transfer.

Conclusions: We show that distributing deep learning models is an effective alternative to sharing patient data. This finding has implications for any collaborative deep learning study.

Publication types

Research Support, N.I.H., Extramural

MeSH terms

Computer Communication Networks
Deep Learning*
Diagnostic Imaging*
Humans
Medical Record Linkage
Neural Networks, Computer

Abstract

Publication types

MeSH terms

Grants and funding