Fetching the paper…
Reading the bibliography…
Online texts with toxic content are a clear threat to the users on social media in particular and society in general.
Semeval-2019 task 6: Identifying and categorizing offensive language in social media (offenseval)
Zampieri, M.; Malmasi, S.; Nakov, P.; Rosenthal, S.; Farra, N.; and Kumar, R. 2019 · 1903
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1907
Earlier work this paper cites.
Zampieri, M.; Nakov, P.; Rosenthal, S.; Atanasova, P.; Karadzhov, G.; Mubarak, H.; Derczynski, L.; Pitenis, Z.; and Çöltekin, Ç. 2020 · 2006
Earlier work this paper cites.
Standardized assessment of reading performance: The new international reading speed texts IReST
Trauzettel-Klosinski, S.; Dietz, K.; Group, I. S.; et al. 2012 · 2012
Earlier work this paper cites.
Generating natural language adversarial examples
Alzantot, M.; Sharma, Y.; Elgohary, A.; Ho, B.-J.; Srivastava, M.; and Chang, K.-W. 2018 · 2018
Earlier work this paper cites.
Black-box generation of adversarial text sequences to evade deep learning classifiers
Gao, J.; Lanchantin, J.; Soffa, M. L.; and Qi, Y. 2018 · 2018
Earlier work this paper cites.
TextBugger: Generating Adversarial Text Against Real-world Applications
Li, J.; Ji, S.; Du, T.; Li, B.; and Wang, T. 2018 · 2018
Earlier work this paper cites.
Semantically equivalent adversarial rules for debugging NLP models
Ribeiro, M. T.; Singh, S.; and Guestrin, C. 2018 · 2018
Earlier work this paper cites.
Interpretable adversarial perturbation in input embedding space for text
Sato, M.; Suzuki, J.; Shindo, H.; and Matsumoto, Y. 2018 · 2018
Earlier work this paper cites.
Semeval-2019 task 5: Multilingual detection of hate speech against immigrants and women in twitter
Basile, V.; Bosco, C.; Fersini, E.; Nozza, D.; Patti, V.; Pardo, F. M. R.; Rosso, P.; and Sanguinetti, M. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Earlier work this paper cites.
Generating natural language adversarial examples through probability weighted word saliency
Ren, S.; Deng, Y.; He, K.; and Che, W. 2019 · 2019
Cited alongside, same era.
Adversarial examples: Attacks and defenses for deep learning
Yuan, X.; He, P.; Zhu, Q.; and Li, X. 2019 · 2019
Cited alongside, same era.
TweetEval:Unified Benchmark and Comparative Evaluation for Tweet Classification
Barbieri, F.; Camacho-Collados, J.; Espinosa-Anke, L.; and Neves, L. 2020 · 2020
Cited alongside, same era.
The FAIR Data principles
FORCE11. 2020 · 2020
Cited alongside, same era.
BAE: BERT-based Adversarial Examples for Text Classification
Garg, S.; and Ramakrishnan, G. 2020 · 2020
Cited alongside, same era.
NeuSpell: A Neural Spelling Correction Toolkit
Jayanthi, S. M.; Pruthi, D.; and Neubig, G. 2020 · 2020
Cited alongside, same era.
Datasheets for datasets
Gebru, T.; Morgenstern, J.; Vecchione, B.; Vaughan, J. W.; Wallach, H.; Iii, H. D.; and Crawford, K. 2021 · 2021
Later among the works it cites.
Hatexplain: A benchmark dataset for explainable hate speech detection
Mathew, B.; Saha, P.; Yimam, S. M.; Biemann, C.; Goyal, P.; and Mukherjee, A. 2021 · 2021
Later among the works it cites.
Openattack: An open-source textual adversarial attack toolkit
Zeng, G.; Qi, F.; Zhou, Q.; Zhang, T.; Hou, B.; Zang, Y.; Liu, Z.; and Sun, M. 2021 · 2021
Later among the works it cites.
Data-Driven Mitigation of Adversarial Text Perturbation
Bhalerao, R.; Al-Rubaie, M.; Bhaskar, A.; and Markov, I. 2022 · 2022
Later among the works it cites.
Perturbations in the Wild: Leveraging Human-Written Text Perturbations for Realistic Adversarial Attack and Defense
Le, T.; Lee, J.; Yen, K.; Hu, Y.; and Lee, D. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Is bert really robust? a strong baseline for natural language attack on text classification and entailment
Jin, D.; Jin, Z.; Zhou, J. T.; and Szolovits, P. 2020 · 2020
Cited alongside, same era.
Bert-attack: Adversarial attack against bert using bert
Li, L.; Ma, R.; Guo, Q.; Xue, X.; and Qiu, X. 2020 · 2020
Cited alongside, same era.
TextAttack: A Framework for Adversarial Attacks, Data Augmentation, and Adversarial Training in NLP
Morris, J.; Lifland, E.; Yoo, J. Y.; Grigsby, J.; Jin, D.; and Qi, Y. 2020 · 2020
Cited alongside, same era.
Word-level textual adversarial attacking as combinatorial optimization
Zang, Y.; Qi, F.; Yang, C.; Liu, Z.; Zhang, M.; Liu, Q.; and Sun, M. 2020 · 2020
Cited alongside, same era.
All that’s’ human’is not gold: Evaluating human evaluation of generated text
Clark, E.; August, T.; Serrano, S.; Haduong, N.; Gururangan, S.; and Smith, N. A. 2021 · 2021
Cited alongside, same era.
Le, T.; Yiran, Y.; Hu, Y.; and Lee, D. 2023 · 2023
Closest in time.
measuring-hate-speech (Revision c65a956)
at UC Berkeley, S. S. D. L. 2024 · 2024
Closest in time.
pyspellchecker
Barrus, T. 2022 · 2025
Closest in time.
Google Spell Check API
Google. 2022 · 2025
Closest in time.
Perspective API
Jigsaw. 2022 · 2025
Closest in time.
Bing Spell Check API
Microsoft. 2022 · 2025
Closest in time.