Loading…

Using explainable machine learning to characterise data drift and detect emergent health risks for emergency department admissions during COVID-19

A key task of emergency departments is to promptly identify patients who require hospital admission. Early identification ensures patient safety and aids organisational planning. Supervised machine learning algorithms can use data describing historical episodes to make ahead-of-time predictions of c...

Full description

Saved in:

Bibliographic Details
Published in:	Scientific reports 2021-11, Vol.11 (1), p.23017-10, Article 23017
Main Authors:	Duckworth, Christopher, Chmiel, Francis P., Burns, Dan K., Zlatev, Zlatko D., White, Neil M., Daniels, Thomas W. V., Kiuber, Michael, Boniface, Michael J.
Format:	Article
Language:	English
Subjects:	639/705/117 692/700 Algorithms Clinical outcomes Coronaviruses COVID-19 COVID-19 - epidemiology Drift Emergency medical care Emergency medical services Emergency Service, Hospital - statistics & numerical data Female Health care Health risk assessment Health risks Hospitalization - statistics & numerical data Humanities and Social Sciences Humans Learning algorithms Machine Learning Male Middle Aged multidisciplinary Pandemics Patient Admission - statistics & numerical data Patient safety Patients Prediction models Public health SARS-CoV-2 - isolation & purification Science Science (multidisciplinary)
Citations:	Items that this one cites Items that cite this one
Online Access:	Get full text
Tags:	Add Tag No Tags, Be the first to tag this record!

Description
Summary:	A key task of emergency departments is to promptly identify patients who require hospital admission. Early identification ensures patient safety and aids organisational planning. Supervised machine learning algorithms can use data describing historical episodes to make ahead-of-time predictions of clinical outcomes. Despite this, clinical settings are dynamic environments and the underlying data distributions characterising episodes can change with time (data drift), and so can the relationship between episode characteristics and associated clinical outcomes (concept drift). Practically this means deployed algorithms must be monitored to ensure their safety. We demonstrate how explainable machine learning can be used to monitor data drift, using the COVID-19 pandemic as a severe example. We present a machine learning classifier trained using (pre-COVID-19) data, to identify patients at high risk of admission during an emergency department attendance. We then evaluate our model’s performance on attendances occurring pre-pandemic (AUROC of 0.856 with 95%CI [0.852, 0.859]) and during the COVID-19 pandemic (AUROC of 0.826 with 95%CI [0.814, 0.837]). We demonstrate two benefits of explainable machine learning (SHAP) for models deployed in healthcare settings: (1) By tracking the variation in a feature’s SHAP value relative to its global importance, a complimentary measure of data drift is found which highlights the need to retrain a predictive model. (2) By observing the relative changes in feature importance emergent health risks can be identified.
ISSN:	2045-2322 2045-2322
DOI:	10.1038/s41598-021-02481-y