Pular para o conteúdo principal

ETHNOS_APP

Início • Busca • Periódicos • Lista 0

Comparing Machine Learning to Regression Methods for Mortality Prediction Using Veterans Affairs Electronic Health Record Clinical Data

Dados Bibliográficos

ID9103213
AutoresBocheng Jing (0000-0002-2021-5055, San Francisco VA Health Care System, autor correspondente), W John Boscardin (0000-0003-3121-9526, San Francisco VA Health Care System, autor correspondente), W James Deardorff (0000-0002-7947-3008, Division of Geriatrics), Sun Young Jeon (San Francisco VA Health Care System, autor correspondente), Alexandra K Lee (0000-0001-9525-3833, San Francisco VA Health Care System, autor correspondente), Anne L Donovan (Anesthesia and Perioperative Medicine, University of California, San Francisco, San Francisco, CA), Sei J Lee (0000-0001-7864-5341, San Francisco VA Health Care System, autor correspondente)
Ano2022
Volume60
Fascículo6
Páginas470-479
Data de publicação2022-06-01
Peer ReviewedSim
Open AccessNão
TipoARTICLE
PeriódicoMedical Care (JOURNAL)
Identificadores do periódicoISSN: 0025-7079 • E-ISSN: 1537-1948
EditoraOvid Technologies (Wolters Kluwer Health) (PUBLISHER)
DOI10.1097/mlr.0000000000001720
PMID35352701
OpenAlexW4220736836
IdiomaEN
Referências citadas34

BACKGROUND: It is unclear whether machine learning methods yield more accurate electronic health record (EHR) prediction models compared with traditional regression methods. OBJECTIVE: The objective of this study was to compare machine learning and traditional regression models for 10-year mortality prediction using EHR data. DESIGN: This was a cohort study. SETTING: Veterans Affairs (VA) EHR data. PARTICIPANTS: Veterans age above 50 with a primary care visit in 2005, divided into separate training and testing cohorts (n= 124,360 each). MEASUREMENTS AND ANALYTIC METHODS: The primary outcome was 10-year all-cause mortality. We considered 924 potential predictors across a wide range of EHR data elements including demographics (3), vital signs (9), medication classes (399), disease diagnoses (293), laboratory results (71), and health care utilization (149). We compared discrimination (c-statistics), calibration metrics, and diagnostic test characteristics (sensitivity, specificity, and positive and negative predictive values) of machine learning and regression models. RESULTS: Our cohort mean age (SD) was 68.2 (10.5), 93.9% were male; 39.4% died within 10 years. Models yielded testing cohort c-statistics between 0.827 and 0.837. Utilizing all 924 predictors, the Gradient Boosting model yielded the highest c-statistic [0.837, 95% confidence interval (CI): 0.835-0.839]. The full (unselected) logistic regression model had the highest c-statistic of regression models (0.833, 95% CI: 0.830-0.835) but showed evidence of overfitting. The discrimination of the stepwise selection logistic model (101 predictors) was similar (0.832, 95% CI: 0.830-0.834) with minimal overfitting. All models were well-calibrated and had similar diagnostic test characteristics. LIMITATION: Our results should be confirmed in non-VA EHRs. CONCLUSION: The differences in c-statistic between the best machine learning model (924-predictor Gradient Boosting) and 101-predictor stepwise logistic models for 10-year mortality prediction were modest, suggesting stepwise regression methods continue to be a reasonable method for VA EHR mortality prediction model development

Cohort · Confidence interval · Logistic regression · Machine learning · Overfitting · Regression · Statistic · Statistics · Stepwise regression · Veterans Affairs · Artificial Intelligence · Artificial Intelligence in Healthcare and Education · Computer Science · Internal Medicine · Machine Learning in Healthcare · Mathematics · Medicine · Sepsis Diagnosis and Treatment

  • Super Learner

    Mark J van der Laan, Eric C Polley et al.•Statistical Applications in…•2007

  • Building Predictive Models in R Using the caret Package

    Open Access•Max Kühn•Journal of Statistical Software•2008

  • Assessing the Performance of Prediction Models

    Ewout W Steyerberg, Andrew J Vickers et al.•Epidemiology•2010

  • Bagging predictors

    Open Access•Leo Breiman•Machine Learning•1996

  • Ranger

    Open Access•Marvin N Wright, Andreas Ziegler•Journal of Statistical Software•2017

  • A working guide to boosted regression trees

    Open Access•Jane Elith, John R Leathwick et al.•Journal of Animal Ecology•2008

  • Regularization Paths for Generalized Linear Models via Coordinate Descent

    Open Access•Jerome Friedman, Jerome H Friedman et al.•Journal of Statistical Software•2010

  • Bagging Predictors

    Open Access•Leo Breiman•Machine Learning•1996

  • Greedy function approximation

    Jerome H Friedman•The Annals of Statistics•2001

  • Random Forests

    Open Access•Leo Breiman•Machine Learning•2001

  • Regression Shrinkage and Selection Via the Lasso

    Open Access•Robert Tibshirani•Journal of the Royal Statistical…•1996

Velocidade de citaçãohistorical
Altamente citadoNão
Ethnos_APP • Projeto Open Source • Licença MIT • Frontend v2.0.0 • Privacidade e Cookies • Documentação da API: api.ethnos.app/docs • Código da API: GitHub • DOI: 10.5281/zenodo.17049435 • Código do Frontend: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae