conftrace_

Papers

20 papers found
HARM: Learning Hate-Aware Reward Model for Evaluating Natural Language Explanations of Offensive Content
Lorenzo Puppi Vecchi, Alceu De Souza Britto Jr., Emerson Cabrera Paraiso et al.
2026 EACL
2026 EACL
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
Brage Eilertsen, Røskva Bjørgfinsdóttir, Francielle Vargas et al.
2026 AAAI
2026 AAAI
X-MuTeST: A Multilingual Benchmark for Explainable Hate Speech Detection and a Novel LLM-Consulted Explanation Framework
Mohammad Zia Ur Rehman, Sai Kartheek Reddy Kasu, Shashivardhan Reddy Koppula et al.
2026 AAAI
SafeLens: Segment-Level Hate Speech Detection in Online Videos
Zhuoran Wang, Dylan Raharja, Yujia Hu et al.
2026 AAAI
2026 ACL
LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
Md Arid Hasan, Firoj Alam, Md Fahad Hossain et al.
2026 ACL
BanHADEX: Towards Explainable HAte Speech Detection in Bangla Using Human Annotated EXplanation
Faisal Hossain Raquib, Akm Moshiur Rahman Mazumder, Md Fahim et al.
2026 ACL
Analyzing Hate Speech Amplification on Fringe Platforms
Anika Ghosh Basu, Humberto Jesus Carlon, Junyi Liu
2026 ACL