Skip to main navigation Skip to search Skip to main content

Recalibrating classifiers for interpretable abusive content detection

  • Bertie Vidgen (Creator)
  • Sam Staton (Creator)
  • Scott Hale (Creator)
  • Ohad Kammar (Creator)
  • Helen Margetts (Creator)
  • Tom Melham (Creator)

Dataset

Description

Dataset and code for the paper, 'Recalibrating classifiers for interpretable abusive content detection' by Vidgen et al. (2020) -- to appear at the NLP + CSS workshop at EMNLP 2020. We provide: 1,000 annotated tweets, sampled using the Davidson classifier with 20 0.05 increments (50 from each) from a dataset of tweets directed against MPs in the UK 2017 General Election 1,000 annotated tweets, sampled using the Perspective classifier with 20 0.05 increments (50 from each) from a dataset of tweets directed against MPs in the UK 2017 General Election Code for recalibration in R and STAN. Annotation guidelines for both datasets. Paper abstract We investigate the use of machine learning classifiers for detecting online abuse in empirical research. We show that uncalibrated classifiers (i.e. where the 'raw' scores are used) align poorly with human evaluations. This limits their use to understand the dynamics, patterns and prevalence of online abuse. We examine two widely used classifiers (created by Perspective and Davidson et al.) on a dataset of tweets directed against candidates in the UK's 2017 general election. A Bayesian approach is presented to recalibrate the raw scores from the classifiers, using probabilistic programming and newly annotated data. We argue that interpretability evaluation and recalibration is integral to the application of abusive content classifiers.

Data Citation

Vidgen, B., Staton, S., Hale, S., Kammar, O., Margetts, H., & Melham, T. (2020). Recalibrating classifiers for interpretable abusive content detection [Data set]. Zenodo. https://doi.org/10.5281/zenodo.4075461
Date made available9 Oct 2020
PublisherZenodo
  • Recalibrating classifiers for interpretable abusive content detection

    Vidgen, B., Staton, S., Hale, S., Kammar, O., Margetts, H., Melham, T. & Szymczak, M., 20 Nov 2020, Proceedings of the Fourth Workshop on Natural Language Processing and Computational Social Science. Bamman, D., Hovy, D., Jurgens, D., O'Connor, B. & Volkova, S. (eds.). ACL Anthology, p. 132-138 7 p.

    Research output: Chapter in Book/Report/Conference proceedingConference contribution

    Open Access
    File

Cite this