0% Complete
Home
/
15th International Conference on Computer and Knowledge Engineering
Enhanced Hate Speech Detection Using Focal Loss and Multi-Head Attention for Imbalanced Social Media Text
Authors :
Ali Rezazadeh
1
Hadi Shahriar Shahhoseini
2
1- Iran University of Science and Technology
2- Iran University of Science and Technology
Keywords :
Hate speech،Cyberbullying،NLP،Natural Language Processing،Deep Learning،Transformer،Attention
Abstract :
Hate speech detection on social media platforms faces significant challenges due to severe class imbalance, where minority classes such as neutral content are substantially underrepresented compared to hate speech and offensive language categories. Traditional deep learning approaches often achieve high overall accuracy by favoring majority classes while failing to adequately detect minority class instances, leading to biased classification systems. This paper presents an enhanced deep learning framework that addresses class imbalance through multiple complementary strategies. Our approach extends DistilBERT with an additional multi-head attention mechanism that refines transformer representations for task-specific semantic understanding. We introduce class-specific processing branches that enable specialized feature learning for each content category, coupled with a comprehensive set of 20 linguistic features capturing domain-specific patterns. To handle extreme class imbalance, we implement an advanced focal loss function with dynamic class weighting and label smoothing, combined with intelligent hybrid sampling strategies and minority class boosting mechanisms. Experimental validation on the Davidson hate speech dataset demonstrates significant improvements over state-of-the-art methods, achieving macro-averaged F1-scores of 91.57\% and weighted-averaged F1-scores of 93.73%. Our approach particularly excels in minority class detection while maintaining robust performance across all categories, with individual class F1-scores ranging from 88.36% to 95.55%. The proposed framework provides a comprehensive solution for imbalanced hate speech classification, combining architectural innovations with advanced loss functions to achieve balanced and effective content moderation capabilities.
Papers List
List of archived papers
Extracting Major Topics of COVID-19 Related Tweets
Faezeh Azizi - Hamed Vahdat-Nejad - Hamideh Hajiabadi - Mohammad Hossein Khosravi
Deep Inside Tor: Exploring Website Fingerprinting Attacks on Tor Traffic in Realistic Settings
Amirhossein Khajehpour - Farid Zandi - Navid Malekghaini - Mahdi Hemmatyar - Naeimeh Omidvar - Mahdi Jafari Siavoshani
GroupRec: Group Recommendation by Numerical Characteristics of Groups in Telegram
Davod Karimpour - Mohammad Ali Zare Chahooki - Ali Hashemi
Security Analysis of MiniApps: Vulnerabilities, Exploits, and a Tailored Mitigation Framework
Keyhan Mohammadi - Arman Moradi - Reza Ebrahimi Atani
A Survey on Semi-Automated and Automated Approaches for Video Annotation
Samin Zare - Mehran Yazdi
AL-YOLO: Accurate and Lightweight Vehicle and Pedestrian Detector in Foggy Weather
Behdad Sadeghian Pour - Hamidreza Mohammadi Jozani - Shahriar Baradaran Shokouhi
Assessing Users' Influence on Respondents in Conversation Quality: A Quantitative Study on Reddit Based on the Cooperative Principle
Afsaneh Habibi - Fattaneh Taghiyareh
Attentional Bi-LSTM for Multivariate Time Series Forecasting on Edge Devices: A Case Study on NanoPi Neo Plus2
Navid Hajizadeh - Saeed Yazdani - Sara Ershadi-Nasab
A Comprehensive Dataset of Real-scene Images for Text Detection and Recognition in Persian
Iman Souzanchi - Ramin Rahimi - Mohammad Ali Majidi Anvari - Atefeh Baniasadi - Ashkan Sadeghi - Mohammad Reza Mohammadi
Ensemble-Based Fraud Detection: A Robust Approach Evaluated on IEEE-CIS
Fatemeh Moradi - Mehran Tarif - Mohammadhossein Homaei
more
Samin Hamayesh - Version 44.5.0