Fetching the paper…
Reading the bibliography…
Current open-domain conversational models can easily be made to talk in inadequate ways.
Build it break it fix it for dialogue safety: Robustness from adversarial human attack
Emily Dinan, Samuel Humeau, Bharath Chintagunta, and Jason Weston. 2019 · 1908
Earlier work this paper cites.
A crowd-based evaluation of abuse response strategies in conversational agents
Amanda Cercas Curry and Verena Rieser. 2019 · 1909
Earlier work this paper cites.
Acute-eval: Improved dialogue evaluation with optimized questions and multi-turn comparisons
Margaret Li, Jason Weston, and Stephen Roller. 2019 · 1909
Earlier work this paper cites.
DialoGPT: Large-scale generative pre-training for conversational response generation
Yizhe Zhang, Siqi Sun, Michel Galley, Yen-Chun Chen, Chris Brockett, Xiang Gao, Jianfeng Gao, Jingjing Liu, and Bill Dolan. 2019 · 1911
Earlier work this paper cites.
Constructive feedback: A key to successful teaching and learning
Martha N Ovando. 1994 · 1994
Earlier work this paper cites.
Towards a human-like open-domain chatbot
Daniel Adiwardana, Minh-Thang Luong, David R So, Jamie Hall, Noah Fiedel, Romal Thoppilan, Zi Yang, Apoorv Kulshreshtha, Gaurav Nemade, Yifeng Lu, et al. 2020 · 2001
Earlier work this paper cites.
Jason Baumgartner, Savvas Zannettou, Brian Keegan, Megan Squire, and Jeremy Blackburn. 2020 · 2001
Earlier work this paper cites.
Reliability in content analysis: Some common misconceptions and recommendations
Klaus Krippendorff. 2004 · 2004
Earlier work this paper cites.
Recipes for building an open-domain chatbot
Stephen Roller, Emily Dinan, Naman Goyal, Da Ju, Mary Williamson, Yinhan Liu, Jing Xu, Myle Ott, Kurt Shuster, Eric M Smith, et al. 2020b · 2004
Earlier work this paper cites.
Open-domain conversational agents: Current progress, open problems, and future directions
Stephen Roller, Y-Lan Boureau, Jason Weston, Antoine Bordes, Emily Dinan, Angela Fan, David Gunning, Da Ju, Margaret Li, Spencer Poff, et al. 2020a · 2006
Earlier work this paper cites.
The psychology of self-defense: Self-affirmation theory
David K Sherman and Geoffrey L Cohen. 2006 · 2006
Earlier work this paper cites.
Neural generation meets real people: Towards emotionally engaging mixed-initiative conversations
Ashwin Paranjape, Abigail See, Kathleen Kenealy, Haojun Li, Amelia Hardy, Peng Qi, Kaushik Ram Sadagopan, Nguyet Minh Phu, Dilara Soylu, and Christopher D Manning. 2020 · 2008
Cited alongside, same era.
ParlAI: A dialog research software platform
Alexander Miller, Will Feng, Dhruv Batra, Antoine Bordes, Adam Fisch, Jiasen Lu, Devi Parikh, and Jason Weston. 2017 · 2017
Cited alongside, same era.
A survey on hate speech detection using natural language processing
Anna Schmidt and Michael Wiegand. 2017 · 2017
Cited alongside, same era.
# metoo: How conversational systems respond to sexual harassment
Amanda Cercas Curry and Verena Rieser. 2018 · 2018
Cited alongside, same era.
Should an agent be ignoring it? a study of verbal abuse types and conversational agents’ response styles
Hyojin Chin and Mun Yong Yi. 2019 · 2019
Cited alongside, same era.
On the dangers of stochastic parrots: Can language models be too big?
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Closest in time.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al. 2021 · 2021
Closest in time.
Countering online hate speech: An nlp perspective
Mudit Chaudhary, Chandni Saxena, and Helen Meng. 2021 · 2021
Closest in time.
Toxicbot: A conversational agent to fight online hate speech
Agustín Manuel de los Riscos and Luis Fernando D’Haro. 2021 · 2021
Closest in time.
Anticipating safety issues in e2e conversational ai: Framework and tooling
Emily Dinan, Gavin Abercrombie, A Stevie Bergman, Shannon Spruit, Dirk Hovy, Y-Lan Boureau, and Verena Rieser. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning from dialogue after deployment: Feed yourself, chatbot!
Braden Hancock, Antoine Bordes, Pierre-Emmanuel Mazare, and Jason Weston. 2019 · 2019
Cited alongside, same era.
Towards empathetic open-domain conversation models: A new benchmark and dataset
Hannah Rashkin, Eric Michael Smith, Margaret Li, and Y-Lan Boureau. 2019 · 2019
Cited alongside, same era.
Empathy is all you need: How a conversational agent should respond to verbal abuse
Hyojin Chin, Lebogang Wame Molefi, and Mun Yong Yi. 2020 · 2020
Cited alongside, same era.
Demographic stability on mechanical turk despite covid-19
Aaron J Moss, Cheskie Rosenzweig, Jonathan Robinson, and Leib Litman. 2020 · 2020
Cited alongside, same era.
Can you put it all together: Evaluating conversational agents’ ability to blend skills
Eric Smith, Mary Williamson, Kurt Shuster, Jason Weston, and Y-Lan Boureau. 2020 · 2020
Cited alongside, same era.
Closest in time.
Towards automatic online hate speech intervention generation using pretrained language model
Raj Ratn Pranesh, Ambesh Shekhar, and Anish Kumar. 2021 · 2021
Closest in time.
Ethical and social risks of harm from language models
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al. 2021 · 2021
Closest in time.
Bot-adversarial dialogue for safe conversational agents
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan. 2021 · 2021
Closest in time.
Generate, prune, select: A pipeline for counterspeech generation against online hate speech
Wanzheng Zhu and Suma Bhat. 2021 · 2021
Closest in time.
Biased assimilation and attitude polarization: The effects of prior theories on subsequently considered evidence
Charles G Lord, Lee Ross, and Mark R Lepper. 1979 · 2098
Closest in time.