2022

Is Your Toxicity My Toxicity? Exploring the Impact of Rater Identity on Toxicity Annotation

Goyal, Nitesh, Kivlichan, Ian, Rosen, Rachel et al.

Understand

Machine learning models are commonly used to detect toxicity in online conversations.

  • These models are trained on datasets annotated by human raters.
  • We explore how raters' self-described identities impact how they annotate toxicity in online comments.
  • We first define the concept of specialized rater pools: rater pools formed based on raters' self-described identities, rather than at random.

Reading the bibliography…