Saltar al contenido principal

ETHNOS_APP

Inicio • Búsqueda • Revistas • Lista 0

Unsupervised natural language processing in the identification of patients with suspected Covid-19 infection

Datos Bibliográficos

ID5223427
AutoresRildo Pinto Da Silva (0000-0001-5718-2747, Universidade de São Paulo), Juliana Tarossi Pollettini (0000-0002-4894-249X, Universidade de São Paulo), Antonio Pazin Filho (0000-0001-5242-329X, Universidade de São Paulo)
Año2023
Volumen39
Número11
Fecha de publicación2023-01-01
Peer ReviewedSí
Open AccessSí
TipoARTICLE
RevistaCadernos de Saude Publica (JOURNAL)
Identificadores de la revistaISSN: 0102-311X • E-ISSN: 1678-4464
EditorialFapUNIFESP (SciELO) (PUBLISHER)
DOI10.1590/0102-311xen243722
OpenAlexW4389317121
SCIELO_PIDS0102-311X2023001105004
IdiomaEN
Referencias citadas30

Patients with post-COVID-19 syndrome benefit from health promotion programs. Their rapid identification is important for the cost-effective use of these programs. Traditional identification techniques perform poorly especially in pandemics. A descriptive observational study was carried out using 105,008 prior authorizations paid by a private health care provider with the application of an unsupervised natural language processing method by topic modeling to identify patients suspected of being infected by COVID-19. A total of 6 models were generated: 3 using the BERTopic algorithm and 3 Word2Vec models. The BERTopic model automatically creates disease groups. In the Word2Vec model, manual analysis of the first 100 cases of each topic was necessary to define the topics related to COVID-19. The BERTopic model with more than 1,000 authorizations per topic without word treatment selected more severe patients - average cost per prior authorizations paid of BRL 10,206 and total expenditure of BRL 20.3 million (5.4%) in 1,987 prior authorizations (1.9%). It had 70% accuracy compared to human analysis and 20% of cases with potential interest, all subject to analysis for inclusion in a health promotion program. It had an important loss of cases when compared to the traditional research model with structured language and identified other groups of diseases - orthopedic, mental and cancer. The BERTopic model served as an exploratory method to be used in case labeling and subsequent application in supervised models. The automatic identification of other diseases raises ethical questions about the treatment of health information by machine learning

Health care · Identification (biology) · Machine learning · Natural language processing · Word2vec · Artificial Intelligence · Artificial Intelligence in Healthcare · Computer Science · COVID-19 diagnosis using AI · Machine Learning in Healthcare · Medicine

  • Long covid—mechanisms, risk factors, and management

    Open Access•Harry Crook, Sanara Raza et al.•BMJ•2021

  • Sentence-Bert

    Open Access•Nils Reimers, Iryna Gurevych•Proceedings of the 2019…•2019

  • Whose life to save? Scarce resources allocation in the Covid-19 outbreak

    Open Access•Chiara Mannelli•Journal of Medical Ethics•2020

  • The Basic Unity of Private Practice and Public Health

    Hugh R Leavell•American Journal of Public Health…•1953

  • A Topic Modeling Comparison Between LDA, NMF, Top2Vec, and Bertopic to Demystify Twitter Posts

    Open Access•Roman Egger, Joanne Yu•Frontiers in Sociology•2022

Velocidad de citaciónhistorical
Altamente citadoNo
Ethnos_APP • Proyecto Open Source • Licencia MIT • Frontend v2.0.0 • Privacidad y Cookies • Documentación de la API: api.ethnos.app/docs • Código de la API: GitHub • DOI: 10.5281/zenodo.17049435 • Código del Frontend: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae