Predicting 1-year mortality of patients with diabetes mellitus in Kazakhstan based on administrative health data using machine learning
© 2023. The Author(s)..
Diabetes mellitus (DM) affects the quality of life and leads to disability, high morbidity, and premature mortality. DM is a risk factor for cardiovascular, neurological, and renal diseases, and places a major burden on healthcare systems globally. Predicting the one-year mortality of patients with DM can considerably help clinicians tailor treatments to patients at risk. In this study, we aimed to show the feasibility of predicting the one-year mortality of DM patients based on administrative health data. We use clinical data for 472,950 patients that were admitted to hospitals across Kazakhstan between mid-2014 to December 2019 and were diagnosed with DM. The data was divided into four yearly-specific cohorts (2016-, 2017-, 2018-, and 2019-cohorts) to predict mortality within a specific year based on clinical and demographic information collected up to the end of the preceding year. We then develop a comprehensive machine learning platform to construct a predictive model of one-year mortality for each year-specific cohort. In particular, the study implements and compares the performance of nine classification rules for predicting the one-year mortality of DM patients. The results show that gradient-boosting ensemble learning methods perform better than other algorithms across all year-specific cohorts while achieving an area under the curve (AUC) between 0.78 and 0.80 on independent test sets. The feature importance analysis conducted by calculating SHAP (SHapley Additive exPlanations) values shows that age, duration of diabetes, hypertension, and sex are the top four most important features for predicting one-year mortality. In conclusion, the results show that it is possible to use machine learning to build accurate predictive models of one-year mortality for DM patients based on administrative health data. In the future, integrating this information with laboratory data or patients' medical history could potentially boost the performance of the predictive models.
Medienart: |
E-Artikel |
---|
Erscheinungsjahr: |
2023 |
---|---|
Erschienen: |
2023 |
Enthalten in: |
Zur Gesamtaufnahme - volume:13 |
---|---|
Enthalten in: |
Scientific reports - 13(2023), 1 vom: 24. Mai, Seite 8412 |
Sprache: |
Englisch |
---|
Beteiligte Personen: |
Alimbayev, Aidar [VerfasserIn] |
---|
Links: |
---|
Themen: |
---|
Anmerkungen: |
Date Completed 26.05.2023 Date Revised 02.06.2023 published: Electronic Citation Status MEDLINE |
---|
doi: |
10.1038/s41598-023-35551-4 |
---|
funding: |
|
---|---|
Förderinstitution / Projekttitel: |
|
PPN (Katalog-ID): |
NLM357281349 |
---|
LEADER | 01000naa a22002652 4500 | ||
---|---|---|---|
001 | NLM357281349 | ||
003 | DE-627 | ||
005 | 20231226210420.0 | ||
007 | cr uuu---uuuuu | ||
008 | 231226s2023 xx |||||o 00| ||eng c | ||
024 | 7 | |a 10.1038/s41598-023-35551-4 |2 doi | |
028 | 5 | 2 | |a pubmed24n1190.xml |
035 | |a (DE-627)NLM357281349 | ||
035 | |a (NLM)37225754 | ||
040 | |a DE-627 |b ger |c DE-627 |e rakwb | ||
041 | |a eng | ||
100 | 1 | |a Alimbayev, Aidar |e verfasserin |4 aut | |
245 | 1 | 0 | |a Predicting 1-year mortality of patients with diabetes mellitus in Kazakhstan based on administrative health data using machine learning |
264 | 1 | |c 2023 | |
336 | |a Text |b txt |2 rdacontent | ||
337 | |a ƒaComputermedien |b c |2 rdamedia | ||
338 | |a ƒa Online-Ressource |b cr |2 rdacarrier | ||
500 | |a Date Completed 26.05.2023 | ||
500 | |a Date Revised 02.06.2023 | ||
500 | |a published: Electronic | ||
500 | |a Citation Status MEDLINE | ||
520 | |a © 2023. The Author(s). | ||
520 | |a Diabetes mellitus (DM) affects the quality of life and leads to disability, high morbidity, and premature mortality. DM is a risk factor for cardiovascular, neurological, and renal diseases, and places a major burden on healthcare systems globally. Predicting the one-year mortality of patients with DM can considerably help clinicians tailor treatments to patients at risk. In this study, we aimed to show the feasibility of predicting the one-year mortality of DM patients based on administrative health data. We use clinical data for 472,950 patients that were admitted to hospitals across Kazakhstan between mid-2014 to December 2019 and were diagnosed with DM. The data was divided into four yearly-specific cohorts (2016-, 2017-, 2018-, and 2019-cohorts) to predict mortality within a specific year based on clinical and demographic information collected up to the end of the preceding year. We then develop a comprehensive machine learning platform to construct a predictive model of one-year mortality for each year-specific cohort. In particular, the study implements and compares the performance of nine classification rules for predicting the one-year mortality of DM patients. The results show that gradient-boosting ensemble learning methods perform better than other algorithms across all year-specific cohorts while achieving an area under the curve (AUC) between 0.78 and 0.80 on independent test sets. The feature importance analysis conducted by calculating SHAP (SHapley Additive exPlanations) values shows that age, duration of diabetes, hypertension, and sex are the top four most important features for predicting one-year mortality. In conclusion, the results show that it is possible to use machine learning to build accurate predictive models of one-year mortality for DM patients based on administrative health data. In the future, integrating this information with laboratory data or patients' medical history could potentially boost the performance of the predictive models | ||
650 | 4 | |a Journal Article | |
650 | 4 | |a Research Support, Non-U.S. Gov't | |
700 | 1 | |a Zhakhina, Gulnur |e verfasserin |4 aut | |
700 | 1 | |a Gusmanov, Arnur |e verfasserin |4 aut | |
700 | 1 | |a Sakko, Yesbolat |e verfasserin |4 aut | |
700 | 1 | |a Yerdessov, Sauran |e verfasserin |4 aut | |
700 | 1 | |a Arupzhanov, Iliyar |e verfasserin |4 aut | |
700 | 1 | |a Kashkynbayev, Ardak |e verfasserin |4 aut | |
700 | 1 | |a Zollanvari, Amin |e verfasserin |4 aut | |
700 | 1 | |a Gaipov, Abduzhappar |e verfasserin |4 aut | |
773 | 0 | 8 | |i Enthalten in |t Scientific reports |d 2011 |g 13(2023), 1 vom: 24. Mai, Seite 8412 |w (DE-627)NLM215703936 |x 2045-2322 |7 nnns |
773 | 1 | 8 | |g volume:13 |g year:2023 |g number:1 |g day:24 |g month:05 |g pages:8412 |
856 | 4 | 0 | |u http://dx.doi.org/10.1038/s41598-023-35551-4 |3 Volltext |
912 | |a GBV_USEFLAG_A | ||
912 | |a GBV_NLM | ||
951 | |a AR | ||
952 | |d 13 |j 2023 |e 1 |b 24 |c 05 |h 8412 |