Capturing Covertly Toxic Speech via Crowdsourcing

Alyssa Lees; Daniel Borkan; Ian Kivlichan; Jorge Nario; Tesh Goyal

2021 EACL EACL 2021

Capturing Covertly Toxic Speech via Crowdsourcing

Abstract

AbstractWe study the task of labeling covert or veiled toxicity in online conversations. Prior research has highlighted the difficulty in creating language models that recognize nuanced toxicity such as microaggressions. Our investigations further underscore the difficulty in parsing such labels reliably from raters via crowdsourcing. We introduce an initial dataset, COVERTTOXICITY, which aims to identify and categorize such comments from a refined rater template. Finally, we fine-tune a comment-domain BERT model to classify covertly offensive comments and compare against existing baselines.

🧭 Keyword Pioneer — covert toxicity

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Natural Language Processing, Security & Privacy

Authors

Alyssa Lees , Daniel Borkan , Ian Kivlichan , Jorge Nario , Tesh Goyal

Topics

Natural Language Processing > Applications > Text Classification

Keywords

bert fine-tuning toxic speech detection covert toxicity

Download PDF

Related papers

Joint Coreference Resolution and Character Linking for Multiparty Conversation 2021

Progressively Pretrained Dense Corpus Index for Open-Domain Question Answering 2021

Crisscrossed Captions: Extended Intramodal and Intermodal Semantic Similarity Judgments for MS-COCO 2021

Representations for Question Answering from Documents with Tables and Text 2021

Gender and Racial Fairness in Depression Research using Social Media 2021