Skip to main content

ETHNOS_APP

Home • Search • Journals • List 0

Missing Data Imputation Using Morphoscopic Traits and Their Performance in the Estimation of Ancestry

Bibliographic Data

ID5284470
AuthorsMichael Kenyhercz, Michael W Kenyhercz (Joint Interoperability Test Command, corresponding author), N V Passalacqua (0000-0003-1634-3844, Western Carolina University), Joseph T Hefner (0000-0001-5535-4410, Michigan State University)
Year2019
Volume2
Issue3
Pages178-188
Publication date2019-11-20
Peer ReviewedYes
Open AccessNo
TypeARTICLE
VenueForensic Anthropology (JOURNAL)
Journal identifiersISSN: 2573-5020 • E-ISSN: 2573-5039
PublisherUniversity Press of Florida (PUBLISHER)
DOI10.5744/fa.2019.1015
OpenAlexW2965988924
LanguageEN
Citations received2

Missing data are an inherent problem in biological anthropology for both reference data sets and individual cases. The goal of data imputation for forensic anthropological applications is to accurately estimate missing values by using other, observed values. To quantify the accuracy of macromorphoscopic data in conditions with slight (10%), moderate (25%), and severe (50%, 75%, and 90%) amounts of missing data, we selected four data-imputation techniques: Hot Deck, iterative robust model-based imputation (IRMI), k-nearest neighbor (k-NN), and the variable medians. Hefner's Macromorphoscopic Databank was used (Hefner 2018); the full sample consisted of 688 individuals from 3 U.S. populations (Blacks, Hispanics, and Whites). Six cranial macromorphoscopic variants were scored in accordance with Hefner (2009). The five data sets with missing data were randomly simulated over multiple iterations (N = 500 each) from the original data. These data sets were compared for agreement using weighted Cohen's kappa and correct classification accuracies over multiple iterations (N = 500) calculated for the original data set. The latter comparisons were also used to examine the effects of imputed data on classification accuracies. Results suggest that IRMI is the most accurate method for imputing missing data, followed by k-NN, in each of the comparisons for nearly all of the variables imputed

Data mining · Data set · Imputation (statistics · Missing data · Pattern recognition (psychology · Statistics · Computer Science · Forensic and Genetic Research · Forensic Anthropology and Bioarchaeology Studies · Mathematics · Race, Genetics, and Society · Artificial Intelligence

  • Estimating inter-individual Mahalanobis distances from mixed incomplete high-dimensional data

    Open Access•H Rathmann, Stephanie Lismann et al.•Journal of Archaeological Science•2023

  • Imputation methods for mixed datasets in bioarchaeology

    Open Access•J Ryan-Despraz, A Wissler•Archaeological and Anthropological…•2024

Unique citing works2
Citations per year0,67
Citation span2023 - 2024 (2)
Citation velocityrecent
Highly citedNo
Citation typesNeutral: 2
Ethnos_APP • Open Source Project • MIT License • Frontend v2.0.0 • Privacy and Cookies • API Documentation: api.ethnos.app/docs • API Source Code: GitHub • DOI: 10.5281/zenodo.17049435 • Frontend Source Code: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae