Fetching the paper…
Reading the bibliography…
Chatbots are used in many applications, e.g., automated agents, smart home assistants, interactive characters in online games, etc.
A Learning Algorithm for Boltzmann Machines
David H. Ackley, Geoffrey E. Hinton, and Terrence J. Sejnowski · 1985
Earlier work this paper cites.
A Neural Probabilistic Language Model
Yoshua Bengio, Réjean Ducharme, and Pascal Vincent · 2000
Earlier work this paper cites.
Unsupervised Modeling of Twitter Conversations
Alan Ritter, Colin Cherry, and Bill Dolan · 2010
Earlier work this paper cites.
Sequence to Sequence Learning with Neural Networks
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le · 2014
Earlier work this paper cites.
Distilling the Knowledge in a Neural Network
Geoffrey E. Hinton, Oriol Vinyals, and Jeffrey Dean · 2015
Earlier work this paper cites.
A Diversity-Promoting Objective Function for Neural Conversation Models
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan · 2016
Earlier work this paper cites.
A Simple, Fast Diverse Decoding Algorithm for Neural Generation
Jiwei Li, Will Monroe, and Dan Jurafsky · 2016
Earlier work this paper cites.
Deep Reinforcement Learning for Dialogue Generation
Jiwei Li, Will Monroe, Alan Ritter, Dan Jurafsky, Michel Galley, and Jianfeng Gao · 2016
Earlier work this paper cites.
Trolls turned Tay, Microsoft’s fun millennial AI bot, into a genocidal maniac
Abby Ohlheiser · 2016
Earlier work this paper cites.
The Limitations of Deep Learning in Adversarial Settings
Nicolas Papernot, Patrick D. McDaniel, Somesh Jha, Matt Fredrikson, Z. Berkay Celik, and Ananthram Swami · 2016
Earlier work this paper cites.
Distillation as a Defense to Adversarial Perturbations Against Deep Neural Networks
Nicolas Papernot, Patrick D. McDaniel, Xi Wu, Somesh Jha, and Ananthram Swami · 2016
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2016
Earlier work this paper cites.
Stealing Machine Learning Models via Prediction APIs
Florian Tramèr, Fan Zhang, Ari Juels, Michael K. Reiter, and Thomas Ristenpart · 2016
Earlier work this paper cites.
Learning End-to-End Goal-Oriented Dialog
Antoine Bordes, Y-Lan Boureau, and Jason Weston · 2017
Earlier work this paper cites.
A Survey on Dialogue Systems: Recent Advances and New Frontiers
Hongshen Chen, Xiaorui Liu, Dawei Yin, and Jiliang Tang · 2017
Earlier work this paper cites.
Automated Hate Speech Detection and the Problem of Offensive Language
Thomas Davidson, Dana Warmsley, Michael W. Macy, and Ingmar Weber · 2017
Earlier work this paper cites.
Beam Search Strategies for Neural Machine Translation
Markus Freitag and Yaser Al-Onaizan · 2017
Earlier work this paper cites.
Kek, Cucks, and God Emperor Trump: A Measurement Study of 4chan’s Politically Incorrect Forum and Its Effects on the Web
Gabriel Emile Hine, Jeremiah Onaolapo, Emiliano De Cristofaro, Nicolas Kourtellis, Ilias Leontiadis, Riginos Samaras, Gianluca Stringhini, and Jeremy Blackburn · 2017
Earlier work this paper cites.
Deceiving Google’s Perspective API Built for Detecting Toxic Comments
Hossein Hosseini, Sreeram Kannan, Baosen Zhang, and Radha Poovendran · 2017
Earlier work this paper cites.
ParlAI: A Dialog Research Software Platform
Alexander H. Miller, Will Feng, Adam Fisch, Jiasen Lu, Dhruv Batra, Antoine Bordes, Devi Parikh, and Jason Weston · 2017
Earlier work this paper cites.
Membership Inference Attacks Against Machine Learning Models
Reza Shokri, Marco Stronati, Congzheng Song, and Vitaly Shmatikov · 2017
Earlier work this paper cites.
Attention is All you Need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Semantics-Enhanced Task-Oriented Dialogue Translation: A Case Study on Hotel Booking
Longyue Wang, Jinhua Du, Liangyou Li, Zhaopeng Tu, Andy Way, and Qun Liu · 2017
Earlier work this paper cites.
Building Task-Oriented Dialogue Systems for Online Shopping
Zhao Yan, Nan Duan, Peng Chen, Ming Zhou, Jianshe Zhou, and Zhoujun Li · 2017
Earlier work this paper cites.
The Web Centipede: Understanding How Web Communities Influence Each Other Through the Lens of Mainstream and Alternative News Sources
Savvas Zannettou, Tristan Caulfield, Emiliano De Cristofaro, Nicolas Kourtellis, Ilias Leontiadis, Michael Sirivianos, Gianluca Stringhini, and Jeremy Blackburn · 2017
Earlier work this paper cites.
Neural Approaches to Conversational AI
Jianfeng Gao, Michel Galley, and Lihong Li · 2018
Earlier work this paper cites.
Advancing the State of the Art in Open Domain Dialog Systems through the Alexa Prize
Chandra Khatri, Behnam Hedayatnia, Anu Venkatesh, Jeff Nunn, Yi Pan, Qing Liu, Han Song, Anna Gottardi, Sanjeev Kwatra, Sanju Pancholi, Ming Cheng, Qinglang Chen, Lauren Stubel, Karthik Gopalakrishnan, Kate Bland, Raefer Gabriel, Arindam Mandal, Dilek Hakkani-Tür, Gene Hwang, Nate Michel, Eric King, and Rohit Prasad · 2018
Cited alongside, same era.
Diverse Beam Search for Improved Description of Complex Scenes
Ashwin K. Vijayakumar, Michael Cogswell, Ramprasaath R. Selvaraju, Qing Sun, Stefan Lee, David J. Crandall, and Dhruv Batra · 2018
Cited alongside, same era.
On the Origins of Memes by Means of Fringe Web Communities
Savvas Zannettou, Tristan Caulfield, Jeremy Blackburn, Emiliano De Cristofaro, Michael Sirivianos, Gianluca Stringhini, and Guillermo Suarez-Tangil · 2018
Cited alongside, same era.
Texygen: A Benchmarking Platform for Text Generation Models
Yaoming Zhu, Sidi Lu, Lei Zheng, Jiaxian Guo, Weinan Zhang, Jun Wang, and Yong Yu · 2018
Cited alongside, same era.
Towards Controllable Biases in Language Generation
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, and Nanyun Peng · 2020
Later among the works it cites.
The Dialogue Dodecathlon: Open-Domain Knowledge and Image Grounded Conversational Agents
Kurt Shuster, Da Ju, Stephen Roller, Emily Dinan, Y-Lan Boureau, and Jason Weston · 2020
Later among the works it cites.
Information Leakage in Embedding Models
Congzheng Song and Ananth Raghunathan · 2020
Later among the works it cites.
MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers
Wenhui Wang, Furu Wei, Li Dong, Hangbo Bao, Nan Yang, and Ming Zhou · 2020
Later among the works it cites.
Recipes for Safety in Open-domain Chatbots
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Emiliano De Cristofaro, Gianluca Stringhini, Athena Vakali, and Nicolas Kourtellis · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Build it Break it Fix it for Dialogue Safety: Robustness from Adversarial Human Attack
Emily Dinan, Samuel Humeau, Bharath Chintagunta, and Jason Weston · 2019
Cited alongside, same era.
Detecting Egregious Responses in Neural Sequence-to-sequence Models
Tianxing He and James R. Glass · 2019
Cited alongside, same era.
Say What I Want: Towards the Dark Side of Neural Dialogue Models
Haochen Liu, Tyler Derr, Zitao Liu, and Jiliang Tang · 2019
Cited alongside, same era.
Language Models are Unsupervised Multitask Learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Cited alongside, same era.
Leveraging Pre-trained Checkpoints for Sequence Generation Tasks
Sascha Rothe, Shashi Narayan, and Aliaksei Severyn · 2019
Cited alongside, same era.
The Risk of Racial Bias in Hate Speech Detection
Maarten Sap, Dallas Card, Saadia Gabriel, Yejin Choi, and Noah A. Smith · 2019
Cited alongside, same era.
Measuring and Characterizing Hate Speech on News Websites
Savvas Zannettou, Mai ElSherief, Elizabeth M. Belding, Shirin Nilizadeh, and Gianluca Stringhini · 2020
Later among the works it cites.
A Quantitative Approach to Understanding Online Antisemitism
Savvas Zannettou, Joel Finkelstein, Barry Bradlyn, and Jeremy Blackburn · 2020
Later among the works it cites.
“Subverting the Jewtocracy”: Online Antisemitism Detection Using Multimodal Deep Learning
Mohit Chandra, Dheeraj Reddy Pailla, Himanshu Bhatia, AadilMehdi J. Sanchawala, Manish Gupta, Manish Shrivastava, and Ponnurangam Kumaraguru · 2021
Later among the works it cites.
“A Virus Has No Religion”: Analyzing Islamophobia on Twitter During the COVID-19 Outbreak
Mohit Chandra, Manvith Reddy, Shradha Sehgal, Saurabh Gupta, Arun Balaji Buduru, and Ponnurangam Kumaraguru · 2021
Later among the works it cites.
Anticipating Safety Issues in E2E Conversational AI: Framework and Tooling
Emily Dinan, Gavin Abercrombie, A. Stevie Bergman, Shannon L. Spruit, Dirk Hovy, Y-Lan Boureau, and Verena Rieser · 2021
Later among the works it cites.
South Korean AI chatbot pulled from Facebook after hate speech towards minorities
Justin McCurry · 2021
Later among the works it cites.
A Survey on Bias and Fairness in Machine Learning
Ninareh Mehrabi, Fred Morstatter, Nripsuta Saxena, Kristina Lerman, and Aram Galstyan · 2021
Later among the works it cites.
Probing Toxic Content in Large Pre-Trained Language Models
Nedjma Ousidhoum, Xinran Zhao, Tianqing Fang, Yangqiu Song, and Dit-Yan Yeung · 2021
Later among the works it cites.
The Evolution of the Manosphere across the Web
Manoel Horta Ribeiro, Jeremy Blackburn, Barry Bradlyn, Emiliano De Cristofaro, Gianluca Stringhini, Summer Long, Stephanie Greenberg, and Savvas Zannettou · 2021
Later among the works it cites.
Recipes for Building an Open-Domain Chatbot
Stephen Roller, Emily Dinan, Naman Goyal, Da Ju, Mary Williamson, Yinhan Liu, Jing Xu, Myle Ott, Eric Michael Smith, Y-Lan Boureau, and Jason Weston · 2021
Later among the works it cites.
On the Safety of Conversational Models: Taxonomy, Dataset, and Benchmark
Hao Sun, Guangxuan Xu, Jiawen Deng, Jiale Cheng, Chujie Zheng, Hao Zhou, Nanyun Peng, Xiaoyan Zhu, and Minlie Huang · 2021
Later among the works it cites.
“Go eat a bat, Chang!”: On the Emergence of Sinophobic Behavior on Web Communities in the Face of COVID-19
Fatemeh Tahmasbi, Leonard Schild, Chen Ling, Jeremy Blackburn, Gianluca Stringhini, Yang Zhang, and Savvas Zannettou · 2021
Later among the works it cites.
SaFeRDialogues: Taking Feedback Gracefully after Conversational Safety Failures
Megan Ung, Jing Xu, and Y-Lan Boureau · 2021
Later among the works it cites.
Challenges in Detoxifying Language Models
Johannes Welbl, Amelia Glaese, Jonathan Uesato, Sumanth Dathathri, John Mellor, Lisa Anne Hendricks, Kirsty Anderson, Pushmeet Kohli, Ben Coppin, and Po-Sen Huang · 2021
Later among the works it cites.
Bot-Adversarial Dialogue for Safe Conversational Agents
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan · 2021
Later among the works it cites.
Racism is a Virus: Anti-Asian Hate and Counterhate in Social Media during the COVID-19 Crisis
Caleb Ziems, Bing He, Sandeep Soni, and Srijan Kumar · 2021
Later among the works it cites.
Robust Conversational Agents against Imperceptible Toxicity Triggers
Ninareh Mehrabi, Ahmad Beirami, Fred Morstatter, and Aram Galstyan · 2022
Closest in time.
The Gospel according to Q: Understanding the QAnon Conspiracy from the Perspective of Canonical Information
Antonis Papasavva, Max Aliapoulios, Cameron Ballard, Emiliano De Cristofaro, Gianluca Stringhini, Savvas Zannettou, and Jeremy Blackburn · 2022
Closest in time.
Red Teaming Language Models with Language Models
Ethan Perez, Saffron Huang, H. Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving · 2022
Closest in time.
Leashing the Inner Demons: Self-Detoxification for Language Models
Canwen Xu, Zexue He, Zhankui He, and Julian J. McAuley · 2022
Closest in time.