Papers
20 papers found
From the Detection of Toxic Spans in Online Discussions to the Analysis of Toxic-to-Civil Transfer
John Pavlopoulos, Leo Laugier, Alexandros Xenos et al.
Mitigating Toxic Degeneration with Empathetic Data: Exploring the Relationship Between Toxicity and Empathy
Allison Lahnala, Charles Welch, Béla Neuendorf et al.
Which One Is More Toxic? Findings from Jigsaw Rate Severity of Toxic Comments
Millon Das, Punyajoy Saha, Mithun Das
Your fairness may vary: Pretrained language model fairness in toxic text classification
Ioana Baldini, Dennis Wei, Karthikeyan Natesan Ramamurthy et al.
NITK-IT_NLP@TamilNLP-ACL2022: Transformer based model for Toxic Span Identification in Tamil
Hariharan LekshmiAmmal, Manikandan Ravikiran, Anand Kumar Madasamy
Detoxifying Language Models with a Toxic Corpus
Yoona Park, Frank Rudzicz
Annotating Targets of Toxic Language at the Span Level
Baran Barbarestani, Isa Maks, Piek Vossen
Detecting Unintended Social Bias in Toxic Language Datasets
Nihar Sahoo, Himanshu Gupta, Pushpak Bhattacharyya
A Stacking-based Efficient Method for Toxic Language Detection on Live Streaming Chat
Yuto Oikawa, Yuki Nakayama, Koji Murakami
Prompt Compression and Contrastive Conditioning for Controllability and Toxicity Reduction in Language Models
David Wingate, Mohammad Shoeybi, Taylor Sorensen
Towards Procedural Fairness: Uncovering Biases in How a Toxic Language Classifier Uses Sentiment Information
Isar Nejadgholi, Esma Balkir, Kathleen Fraser et al.
Detecting Unintended Social Bias in Toxic Language Datasets
Nihar Sahoo, Himanshu Gupta, Pushpak Bhattacharyya
DeToxy: A Large-Scale Multimodal Dataset for Toxicity Classification in Spoken Utterances
Sreyan Ghosh, Samden Lepcha, S Sakshi et al.
Explaining Toxic Text via Knowledge Enhanced Text Generation
Rohit Sridhar, Diyi Yang
Robust Conversational Agents against Imperceptible Toxicity Triggers
Ninareh Mehrabi, Ahmad Beirami, Fred Morstatter et al.
Annotators with Attitudes: How Annotator Beliefs And Identities Bias Toxic Language Detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna et al.
Towards Toxic Positivity Detection
Ishan Sanjeev Upadhyay, KV Aditya Srivatsa, Radhika Mamidi
The subtle language of exclusion: Identifying the Toxic Speech of Trans-exclusionary Radical Feminists
Christina Lu, David Jurgens
Lost in Distillation: A Case Study in Toxicity Modeling
Alyssa Chvasta, Alyssa Lees, Jeffrey Sorensen et al.
A Decision Support System to Predict Acute Fish Toxicity
Anders L Madsen, S. Jannicke Moe, Thomas Braunbeck et al.