Fetching the paper…
Reading the bibliography…
Detecting social bias in text is challenging due to nuance, subjectivity, and difficulty in obtaining good quality labeled datasets at scale, especially given the evolving nature of social biases and society.
Racial bias in hate speech and abusive language detection datasets
Thomas Davidson, Debasmita Bhattacharya, and Ingmar Weber. 2019 · 1905
Earlier work this paper cites.
Megatron-lm: Training multi-billion parameter language models using model parallelism
Mohammad Shoeybi, Mostofa Patwary, Raul Puri, Patrick LeGresley, Jared Casper, and Bryan Catanzaro. 2019 · 1909
Earlier work this paper cites.
Towards non-toxic landscapes: Automatic toxic comment detection using dnn
Ashwin Geet D’sa, Irina Illina, and Dominique Fohr. 2019 · 1911
Earlier work this paper cites.
A joint many-task model: Growing a neural network for multiple nlp tasks
Kazuma Hashimoto, Caiming Xiong, Yoshimasa Tsuruoka, and Richard Socher. 2017 · 1933
Earlier work this paper cites.
Controlling other people: The impact of power on stereotyping
Susan T Fiske. 1993 · 1993
Earlier work this paper cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2005
Earlier work this paper cites.
Hatebert: Retraining bert for abusive language detection in english
Tommaso Caselli, Valerio Basile, Jelena Mitrović, and Michael Granitzer. 2020 · 2010
Earlier work this paper cites.
Thou shalt not discriminate: How emphasizing moral ideals rather than obligations increases whites’ support for social equality
Serena Does, Belle Derks, and Naomi Ellemers. 2011 · 2011
Earlier work this paper cites.
Making pre-trained language models better few-shot learners
Tianyu Gao, Adam Fisch, and Danqi Chen. 2020 · 2012
Earlier work this paper cites.
Automatic identification of personal insults on social news sites
Sara Owsley Sood, Elizabeth F Churchill, and Judd Antin. 2012 · 2012
Earlier work this paper cites.
Hate speech detection with comment embeddings
Nemanja Djuric, Jing Zhou, Robin Morris, Mihajlo Grbovic, Vladan Radosavljevic, and Narayan Bhamidipati. 2015 · 2015
Earlier work this paper cites.
Sarcasm detection on twitter: A behavioral modeling approach
Ashwin Rajadesingan, Reza Zafarani, and Huan Liu. 2015 · 2015
Earlier work this paper cites.
Analyzing biases in human perception of user age and gender from text
Lucie Flekova, Jordan Carpenter, Salvatore Giorgi, Lyle Ungar, and Daniel Preoţiuc-Pietro. 2016 · 2016
Earlier work this paper cites.
Abusive language detection in online user content
Chikashi Nobata, Joel Tetreault, Achint Thomas, Yashar Mehdad, and Yi Chang. 2016 · 2016
Earlier work this paper cites.
Deep multi-task learning with low level tasks supervised at lower layers
Anders Søgaard and Yoav Goldberg. 2016 · 2016
Earlier work this paper cites.
Hateful symbols or hateful people? predictive features for hate speech detection on Twitter
Zeerak Waseem and Dirk Hovy. 2016 · 2016
Earlier work this paper cites.
Deep learning for hate speech detection in tweets
Pinkesh Badjatiya, Shashank Gupta, Manish Gupta, and Vasudeva Varma. 2017 · 2017
Earlier work this paper cites.
A measurement study of hate speech in social media
Mainack Mondal, Leandro Araújo Silva, and Fabrício Benevenuto. 2017 · 2017
Earlier work this paper cites.
Deeper attention to abusive user content moderation
John Pavlopoulos, Prodromos Malakasiotis, and Ion Androutsopoulos. 2017 · 2017
Earlier work this paper cites.
An overview of multi-task learning in deep neural networks
Sebastian Ruder. 2017 · 2017
Earlier work this paper cites.
A web of hate: Tackling hateful speech in online social spaces
Haji Mohammad Saleem, Kelly P Dillon, Susan Benesch, and Derek Ruths. 2017 · 2017
Earlier work this paper cites.
Understanding abuse: A typology of abusive language detection subtasks
Zeerak Waseem, Thomas Davidson, Dana Warmsley, and Ingmar Weber. 2017 · 2017
Cited alongside, same era.
Using machine learning to support qualitative coding in social science: Shifting the focus to ambiguity
Nan-Chen Chen, Margaret Drouhard, Rafal Kocielnik, Jina Suh, and Cecilia R Aragon. 2018 · 2018
Cited alongside, same era.
Learning representations for detecting abusive language
Magnus Sahlgren, Tim Isbister, and Fredrik Olsson. 2018 · 2018
Cited alongside, same era.
Anatomy of online hate: developing a taxonomy and machine learning models for identifying and classifying hate in online news media
Joni Salminen, Hind Almerekhi, Milica Milenković, Soon-gyo Jung, Jisun An, Haewoon Kwak, and Bernard J Jansen. 2018 · 2018
Cited alongside, same era.
Detecting hate speech on twitter using a convolution-gru based deep neural network
Ziqi Zhang, David Robinson, and Jonathan Tepper. 2018 · 2018
Evaluating commonsense in pre-trained language models
Xuhui Zhou, Yue Zhang, Leyang Cui, and Dandan Huang. 2020 · 2020
Later among the works it cites.
Improving gender fairness of pre-trained language models without catastrophic forgetting
Zahra Fatemi, Chen Xing, Wenhao Liu, and Caiming Xiong. 2021 · 2021
Closest in time.
Hatebase
Hatebase.org. 2021 · 2021
Closest in time.
Five sources of bias in natural language processing
Dirk Hovy and Shrimai Prabhumoye. 2021 · 2021
Closest in time.
Cutting down on prompts and parameters: Simple few-shot learning with language models
Robert L Logan IV, Ivana Balažević, Eric Wallace, Fabio Petroni, Sameer Singh, and Sebastian Riedel. 2021 · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
CONAN - COunter NArratives through nichesourcing: a multilingual dataset of responses to fight online hate speech
Yi-Ling Chung, Elizaveta Kuzmenko, Serra Sinem Tekiroglu, and Marco Guerini. 2019 · 2019
Cited alongside, same era.
Commonsense knowledge mining from pretrained models
Joe Davison, Joshua Feldman, and Alexander M Rush. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Overview of the hasoc track at fire 2019: Hate speech and offensive content identification in indo-european languages
Thomas Mandl, Sandip Modha, Prasenjit Majumder, Daksh Patel, Mohana Dave, Chintak Mandlia, and Aditya Patel. 2019 · 2019
Cited alongside, same era.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Cited alongside, same era.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Online hate ratings vary by extremes: A statistical analysis
Joni Salminen, Hind Almerekhi, Ahmed Mohamed Kamel, Soon-gyo Jung, and Bernard J Jansen. 2019 · 2019
Cited alongside, same era.
Sampling: design and analysis
Sharon L Lohr. 2021 · 2021
Closest in time.
Perspective | developers
PerspectiveAPI. 2021 · 2021
Closest in time.
Solid: A large-scale semi-supervised dataset for offensive language identification
Sara Rosenthal, Pepa Atanasova, Georgi Karadzhov, Marcos Zampieri, and Preslav Nakov. 2021 · 2021
Closest in time.
Self-diagnosis and self-debiasing: A proposal for reducing corpus-based bias in nlp
Timo Schick, Sahana Udupa, and Hinrich Schütze. 2021 · 2021
Closest in time.
List of dirty, naughty, obscene, and otherwise bad words
Shutterstock. 2013 · 2021
Closest in time.
Ai and the list of dirty, naughty, obscene, and otherwise bad words | wired
Tom Simonite. 2021 · 2021
Closest in time.
Github - surge-ai/profanity: The world’s largest profanity list
SurgeAI. 2021 · 2021
Closest in time.
Offensive/profane word list
Luis von Ahn. 2021 · 2021
Closest in time.
Ethical and social risks of harm from language models
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al. 2021 · 2021
Closest in time.
Factual probing is [MASK]: Learning vs. learning to recall
Zexuan Zhong, Dan Friedman, and Danqi Chen. 2021 · 2021
Closest in time.
Roc-auc-score
Scikit-learn. 2022a · 2022
Closest in time.
Tfidfvectorizer
Scikit-learn. 2022b · 2022
Closest in time.
Shaden Smith, Mostofa Patwary, Brandon Norick, Patrick LeGresley, Samyam Rajbhandari, Jared Casper, Zhun Liu, Shrimai Prabhumoye, George Zerveas, Vijay Korthikanti, et al. 2022 · 2022
Closest in time.
Learning from noisy labels with deep neural networks: A survey
Hwanjun Song, Minseok Kim, Dongmin Park, Yooju Shin, and Jae-Gil Lee. 2022 · 2022
Closest in time.
Part-of-speech-tagging
POS-tagger Spacy. 2022 · 2022
Closest in time.