Objective: Sharing medical data between institutions is difficult in practice due to data protection laws and official procedures within institutions. Therefore, most existing algorithms are trained on relatively small electroencephalogram (EEG) data sets which is likely to be detrimental to prediction accuracy. In this work, we simulate a case when the data can not be shared by splitting the publicly available data set into disjoint sets representing data in individual institutions.

Methods And Procedures: We propose to train a (local) detector in each institution and aggregate their individual predictions into one final prediction. Four aggregation schemes are compared, namely, the majority vote, the mean, the weighted mean and the Dawid-Skene method. The method was validated on an independent data set using only a subset of EEG channels.

Results: The ensemble reaches accuracy comparable to a single detector trained on all the data when sufficient amount of data is available in each institution.

Conclusion: The weighted mean aggregation scheme showed best performance, it was only marginally outperformed by the Dawid-Skene method when local detectors approach performance of a single detector trained on all available data.

Clinical Impact: Ensemble learning allows training of reliable algorithms for neonatal EEG analysis without a need to share the potentially sensitive EEG data between institutions.

Download full-text PDF

Source
http://www.ncbi.nlm.nih.gov/pmc/articles/PMC9484737PMC
http://dx.doi.org/10.1109/JTEHM.2022.3201167DOI Listing

Publication Analysis

Top Keywords

data
11
ensemble learning
8
data institutions
8
eeg data
8
data set
8
dawid-skene method
8
single detector
8
detector trained
8
learning individual
4
individual neonatal
4

Similar Publications

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!