Unlocking Bias Detection
Leveraging Transformer-Based Models for Content Analysis
Datos Bibliográficos
| ID | 22107366 |
|---|---|
| Autores | Shaina Raza (0000-0003-1061-5845, Vector Institute), Oluwanifemi Bamgbose (Vector Institute), Veronica Chatrath (0000-0003-4790-3963, Vector Institute), Shardule Ghuge (0009-0001-0728-0733, Vector Institute), Yan Sidyakin (Systems, Applications & Products in Data Processing (Canada)), Abdullah Yahya Mohammed Muaad (0000-0001-8304-9261, University of Mysore) |
| Año | 2024 |
| Volumen | 11 |
| Número | 5 |
| Páginas | 6422-6434 |
| Fecha de publicación | 2024-10-01 |
| Peer Reviewed | Sí |
| Open Access | Sí |
| Tipo | ARTICLE |
| Revista | IEEE Transactions on Computational Social Systems (JOURNAL) |
| Identificadores de la revista | ISSN: 2329-924X • E-ISSN: 2373-7476 |
| Editorial | Institute of Electrical and Electronics Engineers (IEEE) (PUBLISHER) |
| DOI | 10.1109/tcss.2024.3392469 |
| OpenAlex | W4398788568 |
| Idioma | EN |
| Referencias citadas | 35 |
Bias detection in text is crucial for combating the spread of negative stereotypes, misinformation, and biased decision-making. Traditional language models frequently face challenges in generalizing beyond their training data and are typically designed for a single task, often focusing on bias detection at the sentence level. To address this, we present the contextualized bi-directional dual transformer (CBDT) classifier. This model combines two complementary transformer networks: the context transformer and the entity transformer, with a focus on improving bias detection capabilities. We have prepared a dataset specifically for training these models to identify and locate biases in texts. Our evaluations across various datasets demonstrate CBDT effectiveness in distinguishing biased narratives from neutral ones and identifying specific biased terms. This work paves the way for applying the CBDT model in various linguistic and cultural contexts, enhancing its utility in bias detection efforts. We also make the annotated dataset available for research purposes
Machine learning · Natural language processing · Sentence · Transformer · Computer Science · Engineering · Natural Language Processing Techniques · Text Readability and Simplification · Topic Modeling · Artificial Intelligence
Persistent Anti-Muslim Bias in Large Language Models
Semantics derived automatically from language corpora contain human-like biases
A Survey on Evaluation of Large Language Models
Fast unfolding of communities in large networks
Bias and Fairness in Large Language Models
A survey of named entity recognition and classification
| Velocidad de citación | historical |
|---|---|
| Altamente citado | No |