Skip to main content

ETHNOS_APP

Home • Search • Journals • List 0

Use of Sequential Hot-Deck Imputation for Missing Health Care Systems Data for Population Health Research

Bibliographic Data

ID9104036
AuthorsElla A Chrenka (0000-0002-3750-7062, HealthPartners Institute, Bloomington, MN, corresponding author), Steven P Dehmer (0000-0003-3508-9239, HealthPartners Institute, Bloomington, MN, corresponding author), Michael V Maciosek (0000-0002-8164-8858, HealthPartners Institute, Bloomington, MN, corresponding author), Inih J Essien (HealthPartners Institute, Bloomington, MN), Inih Essien (0000-0003-0775-8255, H ea lt hP ar tn er s, corresponding author), Bjorn C Westgard (HealthPartners Institute, Bloomington, MN), Bjorn Westgard (0000-0003-1275-753X, Regions Hospital, corresponding author)
Year2024
Volume62
Issue5
Pages319-325
Publication date2024-05-01
Peer ReviewedYes
Open AccessNo
TypeARTICLE
VenueMedical Care (JOURNAL)
Journal identifiersISSN: 0025-7079 • E-ISSN: 1537-1948
PublisherOvid Technologies (Wolters Kluwer Health) (PUBLISHER)
DOI10.1097/mlr.0000000000001995
PMID38546379
OpenAlexW4393253378
LanguageEN
References cited26

Electronic medical record (EMR) data present many opportunities for population health research. The use of EMR data for population risk models can be impeded by the high proportion of missingness in key patient variables. Common approaches like complete case analysis and multiple imputation may not be appropriate for some population health initiatives that require a single, complete analytic data set. In this study, we demonstrate a sequential hot-deck imputation (HDI) procedure to address missingness in a set of cardiometabolic measures in an EMR data set. We assessed the performance of sequential HDI within the individual variables and a commonly used composite risk score. A data set of cardiometabolic measures based on EMR data from 2 large urban hospitals was used to create a benchmark data set with simulated missingness. Sequential HDI was applied, and the resulting data were used to calculate atherosclerotic cardiovascular disease risk scores. The performance of the imputation approach was assessed using a set of metrics to evaluate the distribution and validity of the imputed data. Of the 567,841 patients, 65% had at least 1 missing cardiometabolic measure. Sequential HDI resulted in the distribution of variables and risk scores that reflected those in the simulated data while retaining correlation. When stratified by age and sex, risk scores were plausible and captured patterns expected in the general population. The use of sequential HDI was shown to be a suitable approach to multivariate missingness in EMR data. Sequential HDI could benefit population health research by providing a straightforward, computationally nonintensive approach to missing EMR data that results in a single analytic data set

Data mining · Data set · Environmental health · Imputation (statistics) · Missing data · Multivariate statistics · Population · Statistics · Chronic Disease Management Strategies · Computer Science · Machine Learning in Healthcare · Mathematics · Medical Coding and Health Information · Medicine

  • Heart Disease and Stroke Statistics—2019 Update

    Emelia J Benjamin, Paul Muntner et al.•Circulation•2019

  • Heart Disease and Stroke Statistics—2017 Update

    Emelia J Benjamin, Michael J Blaha et al.•Circulation•2017

  • Imputation with the R Package VIM

    Open Access•Alexander Kowarik, Matthias Templ•Journal of Statistical Software•2016

  • The Impact of eHealth on the Quality and Safety of Health Care

    Open Access•Ashly D Black, Josip Car et al.•PLoS Medicine•2011

  • An overview of clinical decision support systems

    Open Access•Reed T Sutton, David Pincock et al.•npj Digital Medicine•2020

  • Dealing with missing data in a multi-question depression scale

    Open Access•Fiona M Shrive, Heather Stuart et al.•BMC Medical Research Methodology•2006

  • A Review of Hot Deck Imputation for Survey Non‐response

    Open Access•Rebecca R Andridge, Roderick J A Little et al.•International Statistical Review•2010

Citation velocityhistorical
Highly citedNo
Ethnos_APP • Open Source Project • MIT License • Frontend v2.0.0 • Privacy and Cookies • API Documentation: api.ethnos.app/docs • API Source Code: GitHub • DOI: 10.5281/zenodo.17049435 • Frontend Source Code: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae