← 返回资源分享
Predicting the Type and Target of Offensive Posts in Social Media
论文
论文
发布时间2019-02-25
发表NAACL 2019 6 · arXiv:1902.09666
作者:Preslav Nakov,Ritesh Kumar,Marcos Zampieri,Shervin Malmasi,Sara Rosenthal,Noura Farra
详细介绍
As offensive content has become pervasive in social media, there has been
much research in identifying potentially offensive messages. However, previous
work on this topic did not consider the problem as a whole, but rather focused
on detecting very specific types of offensive content, e.g., hate speech,
cyberbulling, or cyber-aggression. In contrast, here we target several
different kinds of offensive content. In particular, we model the task
hierarchically, identifying the type and the target of offensive messages in
social media. For this purpose, we complied the Offensive Language
Identification Dataset (OLID), a new dataset with tweets annotated for
offensive content using a fine-grained three-layer annotation scheme, which we
make publicly available. We discuss the main similarities and differences
between OLID and pre-existing datasets for hate speech identification,
aggression detection, and similar tasks. We further experiment with and we
compare the performance of different machine learning models on OLID.
代码仓库 (2)
joeykay9/offenseval
idontflow/olid
