Skip to main content

ETHNOS_APP

Home • Search • Journals • List 0

Automated Hate Speech Detection and the Problem of Offensive Language

Bibliographic Data

ID23328286
AuthorsThomas R Davidson (0000-0002-5947-7490, Cornell University), Thomas Davidson (0000-0002-9393-2573), Dana Warmsley (Cornell University), Michael W Macy (0000-0003-0024-5027, Cornell University), Michael Macy, Ingmar Weber (0000-0003-4169-2579, Hamad bin Khalifa University)
Year2017
Volume11
Issue1
Pages512-515
Publication date2017-05-03
Peer ReviewedYes
Open AccessYes
TypeARTICLE
VenueProceedings of the International AAAI Conference on Web and Social Media (JOURNAL)
Journal identifiersISSN: 2162-3449 • E-ISSN: 2334-0770
PublisherAssociation for the Advancement of Artificial Intelligence (AAAI) (PUBLISHER)
DOI10.1609/icwsm.v11i1.14955
OpenAlexW2595653137
LanguageEN
Citations received140
References cited4

A key challenge for automatic hate-speech detection on social media is the separation of hate speech from other instances of offensive language. Lexical detection methods tend to have low precision because they classify all messages containing particular terms as hate speech and previous work using supervised learning has failed to distinguish between the two categories. We used a crowd-sourced hate speech lexicon to collect tweets containing hate speech keywords. We use crowd-sourcing to label a sample of these tweets into three categories: those containing hate speech, only offensive language, and those with neither. We train a multi-class classifier to distinguish between these different categories. Close analysis of the predictions and the errors shows when we can reliably separate hate speech from other offensive language and when this differentiation is more difficult. We find that racist and homophobic tweets are more likely to be classified as hate speech but that sexist tweets are generally classified as offensive. Tweets without explicit hate keywords are also more difficult to classify.

Classifier (UML) · Lexicon · Natural language processing · Offensive · Speech processing · Speech recognition · Voice activity detection · Artificial Intelligence · Computer Science · Hate Speech and Cyberbullying Detection · Internet Traffic Analysis and Secure E-voting · Mathematics · Spam and Phishing Detection

  • Language as data, hate as task

    Open Access•Stefanie Ullmann•Discourse Studies•2026

  • ‘Toxic’ memes

    Open Access•Delfina S Martinez-Pandiani, Erik Tjong Kim Sang et al.•Online Social Networks and Media•2025

  • Towards countering hate speech against journalists on social media

    Open Access•Polychronis Charitidis, Stavros Doropoulos et al.•Online Social Networks and Media•2020

  • Gendered hate speech in YouTube and YouNow comments

    Open Access•Nicola Döring, M Rohangis Mohseni•Studies in Communication and Media•2020

  • Constructive Aggression? Multiple Roles of Aggressive Content in Political Discourse on Russian YouTube

    Open Access•Svetlana S Bodrunova, Anna Litvinenko et al.•Media and Communication•2021

  • Measuring Causal Effects of Civil Communication without Randomization

    Open Access•Tony Liu, Lyle Ungar et al.•Proceedings of the International…•2024

  • A Visual Approach to Tracking Emotional Sentiment Dynamics in Social Network Commentaries

    Open Access•Ismail Hossain, Sai Puppala et al.•Proceedings of the International…•2024

  • The Dynamics of Political Incivility on Twitter

    Open Access•Yannis Theocharis, Pablo Barberá Aresté et al.•SAGE Open•2020

  • Hate and Incivilities in Hashtags against Women Candidates in Chile (2021–2022)

    Open Access•José Beltrán, Paula Walker et al.•Social Sciences•2023

  • Exploring the evolution of posting behavior and language use in a racially and ethnically motivated extremist forum

    Open Access•Sydney Litterer, Ryan Scrivens et al.•Behavioral Sciences of Terrorism…•2025

  • Govor mržnje u hrvatskom medijskom prostoru

    Open Access•Marko Poljak, Jelena Hadžić et al.•In medias res•2020

  • A Critical Stylistic Study of Cyber Trolls’ Comments on the Al-Jazeera Arabic TV Channel’s YouTube Videos Concerning the Israeli-Iranian 2024 Conflict

    Open Access•Sadiq Mahdi Kadhim Al Shamiri•Theory and Practice in Language…•2025

  • Polarización, desinformación y expresiones de odio en Twitter. Caso grupos políticos nacionalistas e independentistas en España

    Open Access•Elias Said-Hung, María Adoración Merino Arribas et al.•Observatorio (OBS•2023

  • ArewaAgainstLGBTQ discourse

    Paul Ayodele Onanuga•African Identities•2023

  • Where's the harm? Screening student evaluations of teaching for offensive, threatening or distressing comments

    Open Access•Matthew J Gibson, Justin Luong et al.•Australasian Journal of…•2022

  • On the social and technical challenges of Web search autosuggestion moderation

    Open Access•Timothy J Hazen, Alexandra Olteanu et al.•First Monday•2022

  • Classifying constructive comments

    Open Access•Varada Kolhatkar, Nithum Thain et al.•First Monday•2023

  • Redes, equipos de monitoreo y aplicaciones móvil para combatir los discursos y delitos de odio en Europa

    Open Access•Roberto Moreno-López, Cesar Arroyo López•Revista Latina de Comunicación…•2022

  • A critical reflection on the use of toxicity detection algorithms in proactive content moderation systems

    Open Access•Mark Warner, Angelika Strohmayer et al.•International Journal of…•2025

  • How AI Bots Have Reinforced Gender Bias in Hate Speech

    Open Access•Daniele Battista, Jessica Camargo Molano•ex aequo - Revista da Associação…•2023

  • Racism in tourism reviews

    Open Access•Li Shu, Gang Li et al.•Tourism Management•2020

  • No2Sectarianism

    Open Access•Alexandra A Siegel, Vivienne Badaan•American Political Science Review•2020

  • The Risk of Racial Bias in Hate Speech Detection

    Open Access•Maarten Sap, Dallas Card et al.•Proceedings of the 57th Annual…•2019

  • Social Media and Democracy

    Open Access•Arye L Hillman, Ronen Gradwohl et al.•Social Media and Democracy•2020

  • The Evolution of the Manosphere across the Web

    Open Access•Maria Helena Ribeiro, Manoel Horta Ribeiro et al.•Proceedings of the International…•2021

  • Misinformation, Disinformation, and Online Propaganda

    Open Access•Allison M Gue, Benjamin Lyons•Social Media and Democracy•2020

  • A Survey on Automatic Detection of Hate Speech in Text

    Open Access•Paula Fortuna, Sergio Nunes•ACM Computing Surveys•2019

  • Social Media, Echo Chambers, and Political Polarization

    Open Access•Pablo Barberá Aresté, Pablo Barberá•Social Media and Democracy•2020

  • Whose harm gets detected? A structured review and conceptual framework for misogyny detection and Dari–Pashto marginalization in AI content moderation

    Open Access•Mursal Dawodi, Juergen Pfeffer et al.•AI & Society•2026

  • A comparative analysis of machine learning algorithms for hate speech detection in social media

    Open Access•Esraa Omran, Estabraq Al Tararwah et al.•Online Journal of Communication…•2023

  • The Role of Victim’s Resilience and Self-Esteem in Experiencing Internet Hate

    Open Access•Wiktoria Jdryczka, Wiktoria Jędryczka et al.•International Journal of…•2022

  • Linguistic models of abusive language

    Open Access•Christian Leuprecht, David B Skillicorn et al.•Dynamics of Asymmetric Conflict•2024

  • Toxic language in online incel communities

    Open Access•Björn Pelzer, Lisa Kaati et al.•SN Social Sciences•2021

  • Modeling aggression propagation on social media

    Open Access•Chrysoula Terizi, Despoina Chatzakou et al.•Online Social Networks and Media•2021

  • Understanding Large Language Model Driven Social Bots

    Open Access•Siyu Li, Jin Yang et al.•IEEE Transactions on Computational…•2026

  • HostileNet

    Open Access•Mohit Bhardwaj, Megha Sundriyal et al.•IEEE Transactions on Computational…•2024

  • Adversarial NLP for Social Network Applications

    Open Access•Izzat Alsmadi, Kashif Ahmad et al.•IEEE Transactions on Computational…•2023

  • Zero-Shot Hate to Non-Hate Text Conversion Using Lexical Constraints

    Open Access•Zishan Ahmad, Vinnakota Sai Sujeeth et al.•IEEE Transactions on Computational…•2023

  • HateThaiSent

    Open Access•Krishanu Maity, A S Poornash et al.•IEEE Transactions on Computational…•2024

  • Backdoor Attack and Defense on Deep Learning

    Open Access•Yang Bai, Gaojie Xing et al.•IEEE Transactions on Computational…•2025

  • Sehc

    Open Access•Soumitra Ghosh, Asif Ekbal et al.•IEEE Transactions on Computational…•2023

  • BiCapsHate

    Open Access•Ashraf Kamal, Tarique Anwar et al.•IEEE Transactions on Computational…•2024

  • Detecting Offensive Language Based on Graph Attention Networks and Fusion Features

    Open Access•Zhenxiong Miao, Xingshu Chen et al.•IEEE Transactions on Computational…•2024

  • Model-Agnostic Meta-Learning for Multilingual Hate Speech Detection

    Open Access•Md Rabiul Awal, Roy Ka-Wei Lee et al.•IEEE Transactions on Computational…•2024

  • How can hate narratives be tracked in online environments? A corpus-based study

    Open Access•Manuel Almagro, Carmela Vieites et al.•Frontiers in Communication•2026

  • Covid-19 and Sinophobia

    Open Access•Matt Costello, Nishant Vishwamitra et al.•Cyberpsychology Behavior and…•2023

  • How Online Content Providers Moderate User‐Generated Content to Prevent Harmful Online Communication

    Open Access•Sabine Einwiller, Sora Kim•Policy & Internet•2020

  • Battle for Britain

    Open Access•Samantha North, Lukasz Piwek et al.•Policy & Internet•2021

  • Hebrew offensive language taxonomy and dataset

    Chaya Liebeskind, Natalia Vanetik et al.•Lodz Papers in Pragmatics•2023

  • An integrated explicit and implicit offensive language taxonomy

    Barbara Lewandowska‐tomaszczyk, Anna Bączkowska et al.•Lodz Papers in Pragmatics•2023

  • Offensive language in media discussion forums

    Olga Dontcheva-Navrátilová, Renata Povolná•Lodz Papers in Pragmatics•2023

  • Inequalities and content moderation

    Open Access•Giovanni De Gregorio, Nicole Stremlau•Global Policy•2023

  • A Feature-Based Approach to Assess Hate Speech in User Comments

    Open Access•Liane Reiners, Christian Schemer•Questions de communication•2020

  • From Internet Meme to the Mainstream

    Open Access•Yibing Sun, Varsha Pendyala et al.•Visual Communication Quarterly•2025

  • Negative Feedback Fuels Hate Speech

    Open Access•Hyo-sun Ryu, Jae Kook Lee•Journalism & Mass Communication…•2025

  • The Real Cancel Culture

    David Lynn Painter, Fiona Bown et al.•Howard Journal of Communications•2026

  • Classification of discussants on cyber-violence incidents

    Jing Wang, Xinran Dai•Information Technology and People•2026

  • Hate in Word and Deed

    Open Access•Susann Wiedlitzka, Gabriele Prati et al.•Journal of Quantitative Criminology•2023

  • Hidden Hate

    Kenji Logie, Noah D Cohen et al.•Justice Quarterly•2025

  • Can we predict the Billboard music chart winner? Machine learning prediction based on Twitter artist-fan interactions

    Jihwan Aum, Jisu Kim et al.•Behaviour and Information…•2023

  • Fair compensation of crowdsourcing work

    Joni Salminen, Ahmed Mohamed Sayed Kamel et al.•Behaviour and Information…•2023

  • Expected behavioural effects of alerts to impolite online news commenters

    Open Access•Joel Kiskola, Thomas Olsson et al.•Behaviour and Information…•2025

  • Study on relationship between adversarial texts and language errors

    Rui Xiao, Kuangyi Zhang et al.•Behaviour and Information…•2025

  • Beyond Incivility

    Open Access•Patrícia Rossini•Communication Research•2022

  • Social media content classification and community detection using deep learning and graph analytics

    Open Access•Mohsan Ali, Mehdi Hassan et al.•Technological Forecasting and…•2023

  • Comunicación en redes y discursos de odio en el contexto español

    Open Access•Roberto Moreno-López, Roberto Moreno López et al.•VISUAL REVIEW International…•2022

  • The expression of hate speech against Afro-descendant, Roma, and LGBTQ+ communities in YouTube comments

    Open Access•Paulo Carvalho, Danielle Caled et al.•Journal of Language Aggression…•2024

  • Intersectionality and the gendered discussion around Muslim Canadian politicians on Twitter

    Open Access•A Al-Rawi, Mina Einifar et al.•Journal of Language Aggression…•2024

  • Offensive language in reactions to public figures in polarised discourse online

    Open Access•Maciej Kulik, Katarzyna Budzynska et al.•Journal of Language Aggression…•2025

  • From Insult to Hate Speech

    Open Access•Sünje Paasch‐Colberg, Sünje Paasch-Colberg et al.•Media and Communication•2021

  • Moral Foundations Twitter Corpus

    Open Access•Joe Hoover, Gwenyth Portillo-Wightman et al.•Social Psychological and…•2020

  • Perceptions and Evaluations of Incivility in Public Online Discussions—Insights From Focus Groups With Different Online Actors

    Open Access•Marike Bormann•Frontiers in Political Science•2022

  • Covid-19 Induced Misinformation on YouTube

    Open Access•Viktor Suter, Morteza Shahrezaye et al.•Frontiers in Political Science•2022

  • Hate Speech Directed at Spanish Female Actors

    Open Access•Lucía Tello Díaz, Lizette Martínez Valerio•Social Inclusion•2025

  • Fueling Toxicity? Studying Deceitful Opinion Leaders and Behavioral Changes of Their Followers

    Open Access•Puck Guldemond, Andreu Casas Salleras et al.•Politics and Governance•2022

  • Developing an Incivility Dictionary for German Online Discussions – a Semi-Automated Approach Combining Human and Artificial Knowledge

    Open Access•Anke Stoll, Lena Marie Wilms et al.•Communication Methods and Measures•2023

  • Promoting Hate Speech by Dehumanizing Metaphors of Immigration

    Isabella Gonçalves•Journalism Practice•2023

  • Framing Migration in Southern European Media

    Open Access•Carlos Arcila-Calderón, David Blanco-Herrero et al.•Journalism Practice•2021

  • The Role of Minority Political Groups in the Dissemination of Disinformation. The Case of Spain

    Elias Said-Hung, Marta Sánchez Esparza et al.•Journalism Practice•2024

  • Can (and should) LLMs perform critical discourse analysis

    Emily Barrow DeJeu•Journal of Multicultural Discourses•2024

  • Doxxing to destroy

    Open Access•Briony Anderson•Crime Media Culture An…•2025

  • Who Leaves Malicious Comments on Online News? An Empirical Study in Korea

    Hyunmi Baek, Moonkyoung Jang et al.•Journalism Studies•2022

  • Evolution of negative visual frames of immigrants and refugees in the main media of Southern Europe

    Open Access•Javier J Amores, Carlos Arcila-Calderón et al.•El Profesional de la Informacion•2020

  • Análisis del discurso público y la toxicidad en X

    Open Access•Koldobika Meso Ayerdi, Urko Peña Alonso et al.•Estudios sobre el Mensaje…•2025

  • Fighting Hate Speech, Silencing Drag Queens? Artificial Intelligence in Content Moderation and Risks to LGBTQ Voices Online

    Open Access•Thiago Dias Oliva, Dennys Marcelo Antonialli et al.•Sexuality & Culture•2020

  • Nationally Representative, Locally Misaligned

    Open Access•Paige Bollen, Joe Higton et al.•Political Analysis•2025

  • Online Polarization and Violence in the United States

    Open Access•Dennis Ekwemnachukwu Okeke, Margaret Adutwumwaa Boateng et al.•Social Science Computer Review•2026

  • An Informed Neural Network for Discovering Historical Documentation Assisting the Repatriation of Indigenous Ancestral Human Remains

    Open Access•Abul Bashar, Richi Nayak et al.•Social Science Computer Review•2023

  • Mesurer l’empreinte antisémite sur YouTube

    Open Access•Benjamin Tainturier, Charles de Dampierre et al.•Bulletin of Sociological…•2023

  • “Moral Turn” and the Entextualization of Homosexuals as Pedophiles in Bolsonaro’s Speeches in Congress (2000 to 2018)

    Open Access•Argus Romero Abreu De Morais, Luiz Paulo Moita-Lopes•Alfa•2024

  • “Virada Moral” E Entextualização Do Homossexual Como Pedófilo Em Falas De Bolsonaro No Congresso (2000 a 2018)

    Open Access•Argus Romero Abreu De Morais, Luiz Paulo Moita-Lopes•Alfa•2024

  • Twitch aggression profile

    Open Access•Seung Woo Chae•Information Communication & Society•2026

  • A survey on moral foundation theory and pre-trained language models

    Open Access•Lorenzo Zangari, Candida M Greco et al.•AI & Society•2025

  • The podcast as the centre of young Colombians’ information consumption in the digital sonosphere

    Andrés Barrios Rubio•Radio Journal International…•2025

  • Cross-cutting interaction, inter-party hostility, and partisan identity

    Open Access•Zeyu Lyu•New Media & Society•2023

  • Technology acceptance and transparency demands for toxic language classification – interviews with moderators of public online discussion fora

    Open Access•Lena Katharina Wilms, Katharina Gerl et al.•Human-Computer Interaction•2024

  • Opposition as “A Mould on the Fatherland

    Ilya Sulzhytski•The Journal of Belarusian Studies•2022

  • Vaniraksha

    Open Access•Venkataramana Battula, B Kranthi Kiran•ShodhKosh: Journal of Visual and…•2026

  • Improving Hebrew offensive language classification using LLM-assisted human-in-the-loop annotation

    Chaya Liebeskind, Yael Yefet•Lodz Papers in Pragmatics•2026

  • The Power of Narrative

    Open Access•Rebekka Kesberg, Liza Mügge•Journal of Community & Applied…•2026

  • Hateful Symbols or Hateful People? Predictive Features for Hate Speech Detection on Twitter

    Open Access•Zeerak Waseem, Dirk Hovy•Proceedings of the NAACL Student…•2016

  • Cyber Hate Speech on Twitter

    Open Access•Pete Burnap, Matthew L Williams•Policy & Internet•2015

  • Vader

    Open Access•Cecelia Hutto, Eric Gilbert•Proceedings of the International…•2014

  • Hate Speech

    Jack M Balkin, Samuel Walker•Journal of American History•1995

Unique citing works140
Citations per year17,5
Citation span2018 - 2026 (9)
Citation velocitycurrent
Highly citedYes
Citation typesNeutral: 118

Tools

Open DOIOpen Access
Ethnos_APP • Open Source Project • MIT License • Frontend v2.0.0 • Privacy and Cookies • API Documentation: api.ethnos.app/docs • API Source Code: GitHub • DOI: 10.5281/zenodo.17049435 • Frontend Source Code: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae