Papers
20 papers found
HateXScore: A Metric Suite for Evaluating Reasoning Quality in Hate Speech Explanations
Yujia Hu, Roy Ka-Wei Lee
HARM: Learning Hate-Aware Reward Model for Evaluating Natural Language Explanations of Offensive Content
Lorenzo Puppi Vecchi, Alceu De Souza Britto Jr., Emerson Cabrera Paraiso et al.
RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection
Yejin Lee, Hyeseon An, Yo-Sub Han
When Words Wear Masks: Detecting Malicious Intents and Hostile Impacts of Online Hate Speech
Priyansh Singhal, Piyush Joshi
HACS-TL: Cross-Script Transfer Learning for Hausa Ajami Hate Speech Detection Using Transformer-Based Architecture
Abdulkadir Shehu Bichi, Muqaddar Ali, Prashant Sharma et al.
Language Choice in Nigerian Social Media Hate Speech
Nneoma C Udeze, Rob Voigt
Benchmarking Hate Speech Detection in Azerbaijani with Turkish Cross-Lingual Transfer and Transformer Models
Tural Alizada, Haim Dubossarsky
Shedding the Facades, Connecting the Domains: Detecting Shifting Multimodal Hate Video with Test-Time Adaptation
Jiao Li, Jian Lang, Xikai Tang et al.
MMBERT: Scaled Mixture-of-Experts Multimodal BERT for Robust Chinese Hate Speech Detection Under Cloaking Perturbations
Qiyao Xue, Yuchen Dou, Zheyuan Ryan Shi et al.
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
Brage Eilertsen, Røskva Bjørgfinsdóttir, Francielle Vargas et al.
TRACE: Textual Relevance Augmentation and Contextual Encoding for Multimodal Hate Detection
Girish A. Koushik, Helen Treharne, Aditya Joshi et al.
X-MuTeST: A Multilingual Benchmark for Explainable Hate Speech Detection and a Novel LLM-Consulted Explanation Framework
Mohammad Zia Ur Rehman, Sai Kartheek Reddy Kasu, Shashivardhan Reddy Koppula et al.
SafeLens: Segment-Level Hate Speech Detection in Online Videos
Zhuoran Wang, Dylan Raharja, Yujia Hu et al.
SAGE: Synergistic Adaptive Gating of Experts for Hateful Video Detection
Jie Huang, Xin Liao, Junjie Wang et al.
LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
Md Arid Hasan, Firoj Alam, Md Fahad Hossain et al.
Confident, Calibrated, or Complicit: Safety Alignment and Ideological Bias in LLM Hate Speech Detection
Sanjeevan Selvaganapathy, Mehwish Nasim
BanHADEX: Towards Explainable HAte Speech Detection in Bangla Using Human Annotated EXplanation
Faisal Hossain Raquib, Akm Moshiur Rahman Mazumder, Md Fahim et al.
Analyzing Hate Speech Amplification on Fringe Platforms
Anika Ghosh Basu, Humberto Jesus Carlon, Junyi Liu
Improving Hate Speech Detection by Fusing Textual and User Interaction Representations in Online Communities
Xu Gao, Dong Jing, Kee-hung Lai