0% Complete
Home
/
15th International Conference on Computer and Knowledge Engineering
Enhanced Hate Speech Detection Using Focal Loss and Multi-Head Attention for Imbalanced Social Media Text
Authors :
Ali Rezazadeh
1
Hadi Shahriar Shahhoseini
2
1- Iran University of Science and Technology
2- Iran University of Science and Technology
Keywords :
Hate speech،Cyberbullying،NLP،Natural Language Processing،Deep Learning،Transformer،Attention
Abstract :
Hate speech detection on social media platforms faces significant challenges due to severe class imbalance, where minority classes such as neutral content are substantially underrepresented compared to hate speech and offensive language categories. Traditional deep learning approaches often achieve high overall accuracy by favoring majority classes while failing to adequately detect minority class instances, leading to biased classification systems. This paper presents an enhanced deep learning framework that addresses class imbalance through multiple complementary strategies. Our approach extends DistilBERT with an additional multi-head attention mechanism that refines transformer representations for task-specific semantic understanding. We introduce class-specific processing branches that enable specialized feature learning for each content category, coupled with a comprehensive set of 20 linguistic features capturing domain-specific patterns. To handle extreme class imbalance, we implement an advanced focal loss function with dynamic class weighting and label smoothing, combined with intelligent hybrid sampling strategies and minority class boosting mechanisms. Experimental validation on the Davidson hate speech dataset demonstrates significant improvements over state-of-the-art methods, achieving macro-averaged F1-scores of 91.57\% and weighted-averaged F1-scores of 93.73%. Our approach particularly excels in minority class detection while maintaining robust performance across all categories, with individual class F1-scores ranging from 88.36% to 95.55%. The proposed framework provides a comprehensive solution for imbalanced hate speech classification, combining architectural innovations with advanced loss functions to achieve balanced and effective content moderation capabilities.
Papers List
List of archived papers
TrackMine: Topic Tracking in Model Mining using Genetic Algorithm
Mohammad Sajad Kasaei - Mohammadreza Sharbaf - Afsaneh Fatemi - Bahman Zamani
Deep Learning Feature Extraction for COVID-19 Detection Algorithm using Computerized Tomography Scan
Maisarah Mohd Sufian - Ervin Gubin Moung - Chong Joon Hou - Ali Farzamnia
Adaptive Sliding Window Optimization for Multi-Dimensional Data Streams Using Reinforcement Learning
Abolfazl Zarghani
Balanced Learning with Optimized Extra Trees Classifier for Reliable Lithology Identification in Imbalanced Well Log Data
Ali Daneshpour - Behnam Yousefimehr - Mehdi Ghatee
GAP: Fault tolerance Improvement of Convolutional Neural Networks through GAN-aided Pruning
Pouya Hosseinzadeh - Yasser Sedaghat - Ahad Harati
Farsi Optical Character Recognition Using a Transformer-based Model
Fatemeh Asadi Zeydabadi - Elham Shabaninia - Hossein Nezamabadi-pour - Melika Shojaee
SUBoost: A Novel Boosting-Based Selective Undersampling for handling Imbalanced Data
Nima Rasi Baghmishe - Jafar Tanha - Ehsan Roshan
A Facial Deepfake Detection Approach using CNN-based Models, Swin Transformer and Classifier Fusion
Alireza Honardoost - Mahdie Rahmati - Babak Nasersharif
Token-Based Access Control for Inter-organization Collaboration in Hyperldger Fabric
Parsa Hedayatnia - Mohammad Ata Jalilian - Mohammad Allahbakhsh - Haleh Amintoosi
An Energy-efficient Clustering Method based on Butterfly Optimization Algorithm by Considering the Criterion of Intra-cluster Distances in WSNs
Fariba Saghi Hadi S. Aghdasi
more
Samin Hamayesh - Version 44.5.0