Fetching the paper…
Reading the bibliography…
Dialogue safety problems severely limit the real-world deployment of neural conversational models and have attracted great research interests recently.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Switchboard SWBD-DAMSL shallow-discourse-function annotation coders manual, draft 13
Daniel Jurafsky, Elizabeth Shriberg, and Debra Biasca. 1997 · 1997
Earlier work this paper cites.
Stupid computer! abuse and social identities
Antonella De Angeli, Rollo Carpenter, et al. 2005 · 2005
Earlier work this paper cites.
I hate you! disinhibition with virtual partners
Antonella De Angeli and Sheryl Brahnam. 2008 · 2008
Earlier work this paper cites.
Controlling style in generated dialogue
Eric Michael Smith, Diana Gonzalez-Rico, Emily Dinan, and Y-Lan Boureau. 2020a · 2009
Earlier work this paper cites.
Robustness testing of language understanding in dialog systems
Jiexi Liu, Ryuichi Takanobu, Jiaxin Wen, Dazhen Wan, Weiran Nie, Hongyan Li, Cheng Li, Wei Peng, and Minlie Huang. 2020 · 2012
Earlier work this paper cites.
A comparative study of chatbots and humans
Amit Mittal, Ayushi Agrawal, Ayushi Chouksey, Rachna Shriwas, and Saloni Agrawal. 2016 · 2016
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
Ethical challenges in data-driven dialogue systems
Peter Henderson, Koustuv Sinha, Nicolas Angelard-Gontier, Nan Rosemary Ke, Genevieve Fried, Ryan Lowe, and Joelle Pineau. 2017 · 2017
Earlier work this paper cites.
Parlai: A dialog research software platform
A. H. Miller, W. Feng, A. Fisch, J. Lu, D. Batra, A. Bordes, D. Parikh, and J. Weston. 2017 · 2017
Earlier work this paper cites.
A survey on hate speech detection using natural language processing
Anna Schmidt and Michael Wiegand. 2017 · 2017
Earlier work this paper cites.
Why we should have seen that coming: comments on microsoft’s tay “experiment,” and wider implications
Marty J Wolf, Keith W Miller, and Frances S Grodzinsky. 2017 · 2017
Earlier work this paper cites.
Ex machina: Personal attacks seen at scale
Ellery Wulczyn, Nithum Thain, and Lucas Dixon. 2017 · 2017
Earlier work this paper cites.
Acm code of ethics and professional conduct
ACM Committee on Professional Ethics. 2018 · 2018
Earlier work this paper cites.
Patient and consumer safety risks when using conversational assistants for medical information: an observational study of siri, alexa, and google assistant
Timothy W Bickmore, Ha Trinh, Stefan Olafsson, Teresa K O’Leary, Reza Asadi, Nathaniel M Rickles, and Ricardo Cruz. 2018 · 2018
Earlier work this paper cites.
# metoo alexa: How conversational systems respond to sexual harassment
Amanda Cercas Curry and Verena Rieser. 2018 · 2018
Earlier work this paper cites.
A survey on automatic detection of hate speech in text
Paula Fortuna and Sérgio Nunes. 2018 · 2018
Earlier work this paper cites.
Gender bias in coreference resolution
Rachel Rudinger, Jason Naradowsky, Brian Leonard, and Benjamin Van Durme. 2018 · 2018
Earlier work this paper cites.
From eliza to xiaoice: Challenges and opportunities with social chatbots
Heung-Yeung Shum, Xiaodong He, and Di Li. 2018 · 2018
Earlier work this paper cites.
Learning gender-neutral word embeddings
Jieyu Zhao, Yichao Zhou, Zeyu Li, Wei Wang, and Kai-Wei Chang. 2018 · 2018
Earlier work this paper cites.
An overview of the features of chatbots in mental health: A scoping review
Alaa A Abd-Alrazaq, Mohannad Alajlani, Ali Abdallah Alalwan, Bridgette M Bewick, Peter Gardner, and Mowafa Househ. 2019 · 2019
Earlier work this paper cites.
Racial bias in hate speech and abusive language detection datasets
Thomas Davidson, Debasmita Bhattacharya, and Ingmar Weber. 2019 · 2019
Earlier work this paper cites.
Build it break it fix it for dialogue safety: Robustness from adversarial human attack
Emily Dinan, Samuel Humeau, Bharath Chintagunta, and Jason Weston. 2019 · 2019
Earlier work this paper cites.
Responsible ai: requirements and challenges
Malik Ghallab. 2019 · 2019
Cited alongside, same era.
Exploring social bias in chatbots using stereotype knowledge
Nayeon Lee, Andrea Madotto, and Pascale Fung. 2019 · 2019
Cited alongside, same era.
Thou shalt not hate: Countering online hate speech
Binny Mathew, Punyajoy Saha, Hardik Tharad, Subham Rajgaria, Prajwal Singhania, Suman Kalyan Maity, Pawan Goyal, and Animesh Mukherje. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Cited alongside, same era.
Dreaddit: A Reddit dataset for stress analysis in social media
Elsbeth Turcan and Kathy McKeown. 2019 · 2019
Cited alongside, same era.
Chatbots and conversational agents in mental health: a review of the psychiatric landscape
Aditya Nrusimha Vaidyam, Hannah Wisniewski, John David Halamka, Matcheri S Kashavan, and John Blake Torous. 2019 · 2019
Chatbots RESET A Framework for Governing Responsible Use of Conversational AI in Healthcare In collaboration with Mitsubishi Chemical Holdings Corporation Contents
World Economic Forum. 2020 · 2020
Later among the works it cites.
Recipes for safety in open-domain chatbots
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan. 2020 · 2020
Later among the works it cites.
MedDialog : Large-scale Medical Dialogue Datasets
Guangtao Zeng, Wenmian Yang, Zeqian Ju, Yue Yang, Sicheng Wang, Ruisi Zhang, Meng Zhou, Jiaqi Zeng, Xiangyu Dong, Ruoyu Zhang, Hongchao Fang, and Penghui Zhu. 2020 · 2020
Later among the works it cites.
Dialogpt: Large-scale generative pre-training for conversational response generation
Yizhe Zhang, Siqi Sun, Michel Galley, Yen-Chun Chen, Chris Brockett, Xiang Gao, Jianfeng Gao, Jingjing Liu, and Bill Dolan. 2020 · 2020
Later among the works it cites.
Just say no: Analyzing the stance of neural dialogue generation in offensive contexts
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
TALKDOWN: A corpus for condescension detection in context
Zijian Wang and Christopher Potts. 2019 · 2019
Cited alongside, same era.
Predicting the type and target of offensive posts in social media
Marcos Zampieri, Shervin Malmasi, Preslav Nakov, Sara Rosenthal, Noura Farra, and Ritesh Kumar. 2019 · 2019
Cited alongside, same era.
Effectiveness and safety of using chatbots to improve mental health: systematic review and meta-analysis
Alaa Ali Abd-Alrazaq, Asma Rababeh, Mohannad Alajlani, Bridgette M Bewick, and Mowafa Househ. 2020 · 2020
Cited alongside, same era.
Towards a human-like open-domain chatbot
Daniel Adiwardana, Minh-Thang Luong, David R. So, Jamie Hall, Noah Fiedel, Romal Thoppilan, Zi Yang, Apoorv Kulshreshtha, Gaurav Nemade, Yifeng Lu, and Quoc V. Le. 2020 · 2020
Cited alongside, same era.
Explainable artificial intelligence (xai): Concepts, taxonomies, opportunities and challenges toward responsible ai
Alejandro Barredo Arrieta, Natalia Díaz-Rodríguez, Javier Del Ser, Adrien Bennetot, Siham Tabik, Alberto Barbado, Salvador García, Sergio Gil-López, Daniel Molina, Richard Benjamins, et al. 2020 · 2020
Cited alongside, same era.
The pushshift reddit dataset
Jason Baumgartner, Savvas Zannettou, Brian Keegan, Megan Squire, and Jeremy Blackburn. 2020 · 2020
Cited alongside, same era.
Ashutosh Baheti, Maarten Sap, Alan Ritter, and Mark O. Riedl. 2021 · 2021
Closest in time.
Assessing political prudence of open-domain chatbots
Yejin Bang, Nayeon Lee, Etsuko Ishii, Andrea Madotto, and Pascale Fung. 2021 · 2021
Closest in time.
Plato-2: Towards building an open-domain chatbot via curriculum learning
Siqi Bao, Huang He, Fan Wang, Hua Wu, Haifeng Wang, Wenquan Wu, Zhen Guo, Zhibin Liu, and Xinchao Xu. 2021 · 2021
Closest in time.
Soumya Barikeri, Anne Lauscher, Ivan Vulić, and Goran Glavaš. 2021 · 2021
Closest in time.
On the dangers of stochastic parrots: Can language models be too big?
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Closest in time.
Bold: Dataset and metrics for measuring biases in open-ended language generation
J. Dhamala, Tony Sun, Varun Kumar, Satyapriya Krishna, Yada Pruksachatkun, Kai-Wei Chang, and Rahul Gupta. 2021 · 2021
Closest in time.
Anticipating safety issues in e2e conversational ai: Framework and tooling
Emily Dinan, Gavin Abercrombie, A. Stevie Bergman, Shannon Spruit, Dirk Hovy, Y-Lan Boureau, and Verena Rieser. 2021 · 2021
Closest in time.
Proposal for a regulation of the european parliament and of the council laying down harmonised rules on artificial intelligence (artificial intelligence act) and amending certain union legislative acts
European Commission. 2021 · 2021
Closest in time.
A systematic review of hate speech automatic detection using natural language processing
Md Saroar Jahan and Mourad Oussalah. 2021 · 2021
Closest in time.
Stefano Menini, Alessio Palmero Aprosio, and Sara Tonelli. 2021 · 2021
Closest in time.
StereoSet: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2020 · 2021
Closest in time.
Resources and benchmark corpora for hate speech detection: a systematic review
Fabio Poletto, Valerio Basile, Manuela Sanguinetti, Cristina Bosco, and Viviana Patti. 2021 · 2021
Closest in time.
"nice try, kiddo": Investigating ad hominems in dialogue responses
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, and Nanyun Peng. 2021 · 2021
Closest in time.
PsyQA: A Chinese dataset for generating long counseling text for mental health support
Hao Sun, Zhenru Lin, Chujie Zheng, Siyang Liu, and Minlie Huang. 2021 · 2021
Closest in time.
Context Sensitivity Estimation in Toxicity Detection
Alexandros Xenos, John Pavlopoulos, and Ion Androutsopoulos. 2021 · 2021
Closest in time.
Bot-adversarial dialogue for safe conversational agents
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan. 2021 · 2021
Closest in time.
A taxonomy, data set, and benchmark for detecting and classifying malevolent dialogue responses
Yangjun Zhang, Pengjie Ren, and M. de Rijke. 2021 · 2021
Closest in time.
Can you put it all together: Evaluating conversational agents’ ability to blend skills
Eric Michael Smith, Mary Williamson, Kurt Shuster, Jason Weston, and Y-Lan Boureau. 2020b · 2030
Closest in time.