Fetching the paper…
Reading the bibliography…
In this work, we demonstrate how existing classifiers for identifying toxic comments online fail to generalize to the diverse concerns of Internet users.
An application of hierarchical kappa-type statistics in the assessment of majority agreement among multiple observers
J. R. Landis and G. G. Koch · 1977
Earlier work this paper cites.
The effects of target variables and settting on perceptions of hate speech
G. Cowan and J. Mettrick · 2002
Earlier work this paper cites.
Modeling the detection of textual cyberbullying
K. Dinakar, R. Reichart, and H. Lieberman · 2011
Earlier work this paper cites.
Detecting hate speech on the World Wide Web
W. Warner and J. Hirschberg · 2012
Earlier work this paper cites.
Stfu noob!: predicting crowdsourced decisions on toxic behavior in online games
J. Blackburn and H. Kwak · 2014
Earlier work this paper cites.
Addressing cyber harassment: An overview of hate crimes in cyberspace
D. K. Citron · 2014
Earlier work this paper cites.
We analyzed more than 1 million comments on 4chan. hate speech there has spiked by 40% since 2015
R. Arthur · 2015
Earlier work this paper cites.
Antisocial behavior in online discussion communities
J. Cheng, C. Danescu-Niculescu-Mizil, and J. Leskovec · 2015
Earlier work this paper cites.
Hate speech detection with comment embeddings
N. Djuric, J. Zhou, R. Morris, M. Grbovic, V. Radosavljevic, and N. Bhamidipati · 2015
Earlier work this paper cites.
The emotional impact of cyberbullying: Differences in perceptions and experiences as a function of role
A. M. G. Gualdo, S. C. Hunter, K. Durkin, P. Arnaiz, and J. J. Maquilón · 2015
Earlier work this paper cites.
Exploring cyberbullying and other toxic behavior in team competition online games
H. Kwak, J. Blackburn, and S. Han · 2015
Earlier work this paper cites.
Automatic detection and prevention of cyberbullying
C. Van Hee, E. Lefever, B. Verhoeven, J. Mennes, B. Desmet, G. De Pauw, W. Daelemans, and V. Hoste · 2015
Earlier work this paper cites.
Lgbt parents and social media: Advocacy, privacy, and disclosure during shifting social movements
L. Blackwell, J. Hardy, T. Ammari, T. Veinot, C. Lampe, and S. Schoenebeck · 2016
Earlier work this paper cites.
Online harassment, digital abuse, and cyberstalking in america
Data & Society · 2016
Earlier work this paper cites.
Abusive language detection in online user content
C. Nobata, J. Tetreault, A. Thomas, Y. Mehdad, and Y. Chang · 2016
Earlier work this paper cites.
Characterizations of online harassment: Comparing policies across social media platforms
J. A. Pater, M. K. Kim, E. D. Mynatt, and C. Fiesler · 2016
Earlier work this paper cites.
Automatic detection of Cyberbullying from Twitter
A. Saravanaraj, J. Sheeba, and S. P. Devaneyan · 2016
Earlier work this paper cites.
Classification and its consequences for online harassment: Design insights from heartmob
L. Blackwell, J. Dimond, S. Schoenebeck, and C. Lampe · 2017
Earlier work this paper cites.
The bag of communities: identifying abusive behavior online with preexisting internet data
E. Chandrasekharan, M. Samory, A. Srinivasan, and E. Gilbert · 2017
Cited alongside, same era.
Mean birds: Detecting aggression and bullying on Twitter
D. Chatzakou, N. Kourtellis, J. Blackburn, E. De Cristofaro, G. Stringhini, and A. Vakali · 2017
Cited alongside, same era.
Automated hate speech detection and the problem of offensive language
T. Davidson, D. Warmsley, M. Macy, and I. Weber · 2017
Cited alongside, same era.
Deceiving google’s perspective api built for detecting toxic comments
H. Hosseini, S. Kannan, B. Zhang, and R. Poovendran · 2017
Cited alongside, same era.
Toxic comment classification challenge
Jigsaw · 2017
Cited alongside, same era.
Online harassment 2017
PEW Research Center · 2017
Racial bias in hate speech and abusive language detection datasets
T. Davidson, D. Bhattacharya, and I. Weber · 2019
Later among the works it cites.
The world’s largest structured repository of regionalized, multilingual hate speech
Hatebase · 2019
Later among the works it cites.
A healthier twitter: Progress and more to do
D. Hicks and D. Gasca · 2019
Later among the works it cites.
Our progress on leading the fight against online bullying
Instagram · 2019
Later among the works it cites.
Alphabet-made chrome extension is designed to tune out toxic comments
D. Lee · 2019
Later among the works it cites.
The trauma floor
C. Newton · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Ex machina: Personal attacks seen at scale
E. Wulczyn, N. Thain, and L. Dixon · 2017
Cited alongside, same era.
Upstanding by design: Bystander intervention in cyberbullying
D. DiFranzo, S. H. Taylor, F. Kazerooni, O. D. Wherry, and N. N. Bazarova · 2018
Cited alongside, same era.
Measuring and mitigating unintended bias in text classification
L. Dixon, J. Li, J. Sorensen, N. Thain, and L. Vasserman · 2018
Cited alongside, same era.
Peer to peer hate: Hate speech instigators and their targets
M. ElSherief, S. Nilizadeh, D. Nguyen, G. Vigna, and E. Belding · 2018
Cited alongside, same era.
Large scale crowdsourcing and characterization of twitter abusive behavior
A. M. Founta, C. Djouvas, D. Chatzakou, I. Leontiadis, J. Blackburn, G. Stringhini, A. Vakali, M. Sirivianos, and N. Kourtellis · 2018
Cited alongside, same era.
All you need is “love” evading hate speech detection
T. Gröndahl, L. Pajola, M. Juuti, M. Conti, and N. Asokan · 2018
Cited alongside, same era.
“i just want to feel safe”: A diary study of safety perceptions on social media
E. M. Redmiles, J. Bodford, and L. Blackwell · 2019
Later among the works it cites.
How well do my results generalize? comparing security and privacy survey results from mturk, web, and telephone samples
E. M. Redmiles, S. Kross, and M. L. Mazurek · 2019
Later among the works it cites.
“they don’t leave us alone anywhere we go”: Gender and digital abuse in south asia
N. Sambasivan, A. Batool, N. Ahmed, T. Matthews, K. Thomas, L. S. Gaytán-Lugo, D. Nemer, E. Bursztein, E. Churchill, and S. Consolvo · 2019
Later among the works it cites.
A quantitative approach to understanding online antisemitism
J. Finkelstein, S. Zannettou, B. Bradlyn, and J. Blackburn · 2020
Later among the works it cites.
Characterizing twitter users who engage in adversarial interactions against political candidates
Y. Hua, M. Naaman, and T. Ristenpart · 2020
Later among the works it cites.
Towards measuring adversarial twitter interactions against candidates in the us midterm elections
Y. Hua, T. Ristenpart, and M. Naaman · 2020
Later among the works it cites.
Twitter still failing women over online violence and abuse
A. International · 2020
Later among the works it cites.
Investigating sampling bias in abusive language detection
D. Razo and S. Kübler · 2020
Later among the works it cites.
Investigating annotator bias with a graph-based approach
M. Wich, H. Al Kuwatly, and G. Groh · 2020
Later among the works it cites.
The disagreement deconvolution: Bringing machine learning performance metrics in line with reality
M. L. Gordon, K. Zhou, K. Patel, T. Hashimoto, and M. S. Bernstein · 2021
Closest in time.
Sok: Hate, harassment, and the changing landscape of online abuse
K. Thomas, D. Akhawe, M. Bailey, D. Boneh, E. Bursztein, S. Consolvo, N. Dell, Z. Durumeric, P. G. Kelley, D. Kumar, D. McCoy, S. Meiklejohn, T. Ristenpart, and G. Stringhini · 2021
Closest in time.
United states census bureau
US Census · 2021
Closest in time.