A Benchmark Dataset for Learning to Intervene in Online Hate Speech

Qian, Jing; Bethke, Anna; Liu, Yinyin; Belding, Elizabeth; Wang, William Yang

Computer Science > Computation and Language

arXiv:1909.04251 (cs)

[Submitted on 10 Sep 2019]

Title:A Benchmark Dataset for Learning to Intervene in Online Hate Speech

Authors:Jing Qian, Anna Bethke, Yinyin Liu, Elizabeth Belding, William Yang Wang

View PDF

Abstract:Countering online hate speech is a critical yet challenging task, but one which can be aided by the use of Natural Language Processing (NLP) techniques. Previous research has primarily focused on the development of NLP methods to automatically and effectively detect online hate speech while disregarding further action needed to calm and discourage individuals from using hate speech in the future. In addition, most existing hate speech datasets treat each post as an isolated instance, ignoring the conversational context. In this paper, we propose a novel task of generative hate speech intervention, where the goal is to automatically generate responses to intervene during online conversations that contain hate speech. As a part of this work, we introduce two fully-labeled large-scale hate speech intervention datasets collected from Gab and Reddit. These datasets provide conversation segments, hate speech labels, as well as intervention responses written by Mechanical Turk Workers. In this paper, we also analyze the datasets to understand the common intervention strategies and explore the performance of common automatic response generation methods on these new datasets to provide a benchmark for future research.

Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Computers and Society (cs.CY)
Cite as:	arXiv:1909.04251 [cs.CL]
	(or arXiv:1909.04251v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1909.04251

Submission history

From: Jing Qian [view email]
[v1] Tue, 10 Sep 2019 03:00:58 UTC (473 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2019-09

Change to browse by:

cs
cs.AI
cs.CY

References & Citations

DBLP - CS Bibliography

listing | bibtex

Jing Qian
Yinyin Liu
Elizabeth M. Belding
William Yang Wang

export BibTeX citation

Computer Science > Computation and Language

Title:A Benchmark Dataset for Learning to Intervene in Online Hate Speech

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:A Benchmark Dataset for Learning to Intervene in Online Hate Speech

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators