UPDF AI

Is ChatGPT better than Human Annotators? Potential and Limitations of ChatGPT in Explaining Implicit Hate Speech

Fan Huang,Haewoon Kwak,Jisun An

2023 · DOI: 10.1145/3543873.3587368
The Web Conference · 316 Citations

TLDR

This work examines whether ChatGPT can be used for providing natural language explanations (NLEs) for implicit hateful speech detection, and design and conduct user studies to evaluate their qualities by comparison with human-written NLEs.

Abstract

Recent studies have alarmed that many online hate speeches are implicit. With its subtle nature, the explainability of the detection of such hateful speech has been a challenging problem. In this work, we examine whether ChatGPT can be used for providing natural language explanations (NLEs) for implicit hateful speech detection. We design our prompt to elicit concise ChatGPT-generated NLEs and conduct user studies to evaluate their qualities by comparison with human-written NLEs. We discuss the potential and limitations of ChatGPT in the context of implicit hateful speech research.