Fetching the paper…
Reading the bibliography…
The emergence of pretrained large language models has led to the deployment of a range of social chatbots for chitchat.
Ensemble-based deep reinforcement learning for chatbots
Heriberto Cuayáhuitl, Donghyeon Lee, Seonghan Ryu, Yongjin Cho, Sungja Choi, Satish Reddy Indurthi, Seunghak Yu, Hyungtak Choi, Inchul Hwang, and Jihie Kim. 2019 · 1908
Earlier work this paper cites.
CTRL: A conditional transformer language model for controllable generation
Nitish Shirish Keskar, Bryan McCann, Lav R. Varshney, Caiming Xiong, and Richard Socher. 2019 · 1909
Earlier work this paper cites.
ACUTE-EVAL: improved dialogue evaluation with optimized questions and multi-turn comparisons
Margaret Li, Jason Weston, and Stephen Roller. 2019 · 1909
Earlier work this paper cites.
Plug and play language models: A simple approach to controlled text generation
Sumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung, Eric Frank, Piero Molino, Jason Yosinski, and Rosanne Liu. 2019 · 1912
Earlier work this paper cites.
Eliza—a computer program for the study of natural language communication between man and machine
Joseph Weizenbaum. 1966 · 1966
Earlier work this paper cites.
The levenberg-marquardt algorithm: implementation and theory
Jorge J Moré. 2006 · 1977
Earlier work this paper cites.
Lessons from a restricted turing test
Stuart M. Shieber. 1994 · 1994
Earlier work this paper cites.
Towards a human-like open-domain chatbot
Daniel Adiwardana, Minh-Thang Luong, David R. So, Jamie Hall, Noah Fiedel, Romal Thoppilan, Zi Yang, Apoorv Kulshreshtha, Gaurav Nemade, Yifeng Lu, and Quoc V. Le. 2020 · 2001
Earlier work this paper cites.
Artificial intelligence, values and alignment
Iason Gabriel. 2020 · 2001
Earlier work this paper cites.
Pre-trained models for natural language processing: A survey
Xipeng Qiu, Tianxiang Sun, Yige Xu, Yunfan Shao, Ning Dai, and Xuanjing Huang. 2020 · 2003
Earlier work this paper cites.
PLATO-2: towards building an open-domain chatbot via curriculum learning
Siqi Bao, Huang He, Fan Wang, Hua Wu, Haifeng Wang, Wenquan Wu, Zhen Guo, Zhibin Liu, and Xinchao Xu. 2020 · 2006
Earlier work this paper cites.
Recipes for safety in open-domain chatbots
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan. 2020 · 2010
Earlier work this paper cites.
Towards an open-domain conversational system fully based on natural language processing
Ryuichiro Higashinaka, Kenji Imamura, Toyomi Meguro, Chiaki Miyazaki, Nozomi Kobayashi, Hiroaki Sugiyama, Toru Hirano, Toshiro Makino, and Yoshihiro Matsuo. 2014 · 2014
Earlier work this paper cites.
A neural network approach to context-sensitive generation of conversational responses
Alessandro Sordoni, Michel Galley, Michael Auli, Chris Brockett, Yangfeng Ji, Margaret Mitchell, Jian-Yun Nie, Jianfeng Gao, and Bill Dolan. 2015 · 2015
Earlier work this paper cites.
A diversity-promoting objective function for neural conversation models
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan. 2016 · 2016
Earlier work this paper cites.
Constrained policy optimization
Joshua Achiam, David Held, Aviv Tamar, and Pieter Abbeel. 2017 · 2017
Cited alongside, same era.
Thinking fast and slow with deep learning and tree search
Thomas Anthony, Zheng Tian, and David Barber. 2017 · 2017
Cited alongside, same era.
A survey on dialogue systems: Recent advances and new frontiers
Hongshen Chen, Xiaorui Liu, Dawei Yin, and Jiliang Tang. 2017 · 2017
Cited alongside, same era.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017 · 2017
Cited alongside, same era.
Legalbot: A deep learning-based conversational agent in the legal domain
Adebayo Kolawole John, Luigi Di Caro, Livio Robaldo, and Guido Boella. 2017 · 2017
Cited alongside, same era.
Towards unified dialogue system evaluation: A comprehensive analysis of current evaluation protocols
Sarah E. Finch and Jinho D. Choi. 2020 · 2020
Later among the works it cites.
Learning to summarize with human feedback
Nisan Stiennon, Long Ouyang, Jeffrey Wu, Daniel Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul F Christiano. 2020 · 2020
Later among the works it cites.
A short survey of pre-trained language models for conversational ai-a new age in nlp
Munazza Zaib, Quan Z. Sheng, and Wei Emma Zhang. 2020 · 2020
Later among the works it cites.
A general language assistant as a laboratory for alignment
Amanda Askell, Yuntao Bai, Anna Chen, Dawn Drain, Deep Ganguli, Tom Henighan, Andy Jones, Nicholas Joseph, Benjamin Mann, Nova DasSarma, Nelson Elhage, Zac Hatfield-Dodds, Danny Hernandez, Jackson Kernion, Kamal Ndousse, Catherine Olsson, Dario Amodei, Tom B. Brown, Jack Clark, Sam McCandlish, Chris Olah, and Jared Kaplan. 2021 · 2021
Later among the works it cites.
An intelligent chatbot using deep learning with bidirectional rnn and attention model
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nicole M. Radziwill and Morgan C. Benton. 2017 · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Cited alongside, same era.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, Timothy P. Lillicrap, Karen Simonyan, and Demis Hassabis. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Context-aware sequence-to-sequence models for conversational systems
Silje Christensen, Simen Johnsrud, Massimiliano Ruocco, and Heri Ramampiaro. 2018 · 2018
Cited alongside, same era.
Scalable agent alignment via reward modeling: a research direction
Jan Leike, David Krueger, Tom Everitt, Miljan Martic, Vishal Maini, and Shane Legg. 2018 · 2018
Cited alongside, same era.
Textual entailment using machine translation evaluation metrics
Tanik Saikh, Sudip Kumar Naskar, Asif Ekbal, and Sivaji Bandyopadhyay. 2018 · 2018
Cited alongside, same era.
Manyu Dhyani and Rajiv Kumar. 2021 · 2021
Later among the works it cites.
Proceedings of the 3rd Workshop on Natural Language Processing for Conversational AI . Association for Computational Linguistics, Online
Alexandros Papangelis, Paweł Budzianowski, Bing Liu, Elnaz Nouri, Abhinav Rastogi, and Yun-Nung Chen, editors. 2021 · 2021
Later among the works it cites.
Recipes for building an open-domain chatbot
Stephen Roller, Emily Dinan, Naman Goyal, Da Ju, Mary Williamson, Yinhan Liu, Jing Xu, Myle Ott, Eric Michael Smith, Y-Lan Boureau, and Jason Weston. 2021 · 2021
Later among the works it cites.
GPT-J-6B: A 6 Billion Parameter Autoregressive Language Model
Ben Wang and Aran Komatsuzaki. 2021 · 2021
Later among the works it cites.
Grounding in social media: An approach to building a chit-chat dialogue model
Ritvik Choudhary and Daisuke Kawahara. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Later among the works it cites.
Deep learning for dialogue systems: Chit-chat and beyond
Rui Yan, Juntao Li, Zhou Yu, et al. 2022 · 2022
Later among the works it cites.
ChatMatch: Evaluating chatbots by autonomous chat tournaments
Ruolan Yang, Zitong Li, Haifeng Tang, and Kenny Zhu. 2022 · 2022
Later among the works it cites.
UniDS: A unified dialogue system for chit-chat and task-oriented dialogues
Xinyan Zhao, Bin He, Yasheng Wang, Yitong Li, Fei Mi, Yajiao Liu, Xin Jiang, Qun Liu, and Huanhuan Chen. 2022 · 2022
Later among the works it cites.
A simple survey of pre-trained language models
Zhenyi Zhu. 2022 · 2022
Later among the works it cites.