Fetching the paper…
Reading the bibliography…
This survey provides a comprehensive review of research on multi-turn dialogue systems, with a particular focus on multi-turn dialogue systems based on large language models (LLMs).
Eliza—a computer program for the study of natural language communication between man and machine
Joseph Weizenbaum · 1966
Earlier work this paper cites.
Artificial paranoia
Kenneth Mark Colby, Sylvia Weber, and Franklin Dennis Hilf · 1971
Earlier work this paper cites.
A form-based dialogue manager for spoken language applications
David Goddeau et al · 1996
Earlier work this paper cites.
Building applied natural language generation systems
EHUD REITER and ROBERT DALE · 1997
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni et al · 2002
Earlier work this paper cites.
Use of kernel deep convex networks and end-to-end learning for spoken language understanding
Li Deng et al · 2012
Earlier work this paper cites.
Towards deeper understanding: Deep convex networks for semantic utterance classification
Gokhan Tur et al · 2012
Earlier work this paper cites.
Dialogue system: A brief review
Suket Arora, Kamaljeet Batra, and Sarabjit Singh · 2013
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
Earlier work this paper cites.
Word-based dialog state tracking with recurrent neural networks
Matthew Henderson, Blaise Thomson, and Steve Young · 2014
Earlier work this paper cites.
Stochastic language generation in dialogue using recurrent neural networks with convolutional sentence reranking
Tsung-Hsien Wen et al · 2015
Earlier work this paper cites.
Semantically conditioned LSTM-based natural language generation for spoken dialogue systems
Tsung-Hsien Wen et al · 2015
Earlier work this paper cites.
Oriol Vinyals and Quoc Le · 2015
Earlier work this paper cites.
Neural responding machine for short-text conversation
Lifeng Shang, Zhengdong Lu, and Hang Li · 2015
Earlier work this paper cites.
A neural network approach to context-sensitive generation of conversational responses
Alessandro Sordoni et al · 2015
Earlier work this paper cites.
Towards end-to-end learning for dialog state tracking and management using deep reinforcement learning
Tiancheng Zhao and Maxine Eskenazi · 2016
Earlier work this paper cites.
Building end-to-end dialogue systems using generative hierarchical neural network models
Iulian Serban et al · 2016
Earlier work this paper cites.
Context-aware natural language generation for spoken dialogue systems
Hao Zhou, Minlie Huang, and Xiaoyan Zhu · 2016
Earlier work this paper cites.
Learning end-to-end goal-oriented dialog
Antoine Bordes, Y-Lan Boureau, and Jason Weston · 2016
Earlier work this paper cites.
Docchat: An information retrieval approach for chatbot engines using unstructured documents
Zhao Yan et al · 2016
Earlier work this paper cites.
A diversity-promoting objective function for neural conversation models
Jiwei et al. Li · 2016
Earlier work this paper cites.
Sequential matching network: A new architecture for multi-turn response selection in retrieval-based chatbots
Yu Wu et al · 2017
Earlier work this paper cites.
A survey on dialogue systems: Recent advances and new frontiers
Hongshen Chen et al · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani et al · 2017
Earlier work this paper cites.
Efficient natural language response suggestion for smart reply
Matthew Henderson et al · 2017
Earlier work this paper cites.
Neural belief tracker: Data-driven dialogue state tracking
Nikola Mrkšić et al · 2017
Earlier work this paper cites.
A network-based end-to-end trainable task-oriented dialogue system
Tsung-Hsien Wen et al · 2017
Earlier work this paper cites.
Key-value retrieval networks for task-oriented dialogue
Mihail et al. Eric · 2017
Earlier work this paper cites.
DailyDialog: A manually labelled multi-turn dialogue dataset
Yanran Li et al · 2017
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford et al · 2018
Earlier work this paper cites.
Global-locally self-attentive encoder for dialogue state tracking
Victor Zhong, Caiming Xiong, and Richard Socher · 2018
Earlier work this paper cites.
MultiWOZ - a large-scale multi-domain Wizard-of-Oz dataset for task-oriented dialogue modelling
Paweł Budzianowski et al · 2018
Earlier work this paper cites.
Personalizing dialogue agents: I have a dog, do you have pets too?
Saizheng Zhang et al · 2018
Earlier work this paper cites.
TripleNet: Triple attention network for multi-turn response selection in retrieval-based chatbots
Wentao Ma et al · 2019
Earlier work this paper cites.
Are training samples correlated? learning to generate dialogue responses with multiple references
Lisong Qiu et al · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford et al · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin et al · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu et al · 2019
Earlier work this paper cites.
Unified language model pre-training for natural language understanding and generation
Li Dong et al · 2019
Earlier work this paper cites.
Mike Lewis et al · 2019
Earlier work this paper cites.
Parameter-efficient transfer learning for nlp
Neil Houlsby et al · 2019
Earlier work this paper cites.
Freelb: Enhanced adversarial training for natural language understanding
Chen Zhu et al · 2019
Earlier work this paper cites.
SUMBT: Slot-utterance matching for universal and scalable belief tracking
Hwaran Lee, Jinsik Lee, and Tae-Yoon Kim · 2019
Earlier work this paper cites.
Bert for joint intent classification and slot filling. arxiv
Qian Chen, Zhu Zhuo, and Wen Wang · 2019
Earlier work this paper cites.
Scalable neural dialogue state tracking
Vevake Balaraman and Bernardo Magnini · 2019
Earlier work this paper cites.
Semantically conditioned dialog response generation via hierarchical disentangled self-attention
Wenhu Chen et al · 2019
Earlier work this paper cites.
Hello, it’s GPT-2 - how can I help you? towards the use of pretrained language models for task-oriented dialogue systems
Paweł Budzianowski and Ivan Vulić · 2019
Earlier work this paper cites.
Multi-hop selector network for multi-turn response selection in retrieval-based chatbots
Chunyuan Yuan et al · 2019
Earlier work this paper cites.
One time of interaction may not be enough: Go deep with an interaction-over-interaction network for response selection in dialogues
Chongyang Tao et al · 2019
Earlier work this paper cites.
Personalizing dialogue agents via meta-learning
Andrea Madotto et al · 2019
Earlier work this paper cites.
Latent retrieval for weakly supervised open domain question answering
Kenton Lee, Ming-Wei Chang, and Kristina Toutanova · 2019
Earlier work this paper cites.
Transferable multi-domain state generator for task-oriented dialogue systems
Chien-Sheng Wu et al · 2019
Earlier work this paper cites.
Persuasion for good: Towards a personalized persuasive dialogue system for social good
Xuewei Wang et al · 2019
Earlier work this paper cites.
Personalized dialogue generation with diversified traits
Yinhe Zheng et al · 2019
Earlier work this paper cites.
Amalgamating knowledge from two teachers for task-oriented dialogue system with adversarial training
Wanwei He et al · 2020
Earlier work this paper cites.
Language models as few-shot learner for task-oriented dialogue systems
Andrea Madotto, Zihan Liu, Zhaojiang Lin, and Pascale Fung · 2020
Earlier work this paper cites.
Scaling laws for neural language models
Jared Kaplan et al · 2020
Earlier work this paper cites.
Pre-trained models for natural language processing: A survey
Xipeng Qiu et al · 2020
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown et al · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel et al · 2020
Earlier work this paper cites.
AdapterHub: A framework for adapting transformers
Jonas Pfeiffer et al · 2020
Earlier work this paper cites.
Revisiting few-sample bert fine-tuning
Tianyi Zhang et al · 2020
Cited alongside, same era.
Better fine-tuning by reducing representational collapse
Armen Aghajanyan et al · 2020
Cited alongside, same era.
Fine-tuning pretrained language models: Weight initializations, data orders, and early stopping
Jesse Dodge, Gabriel Ilharco, Roy Schwartz, Ali Farhadi, Hannaneh Hajishirzi, and Noah A Smith · 2020
Cited alongside, same era.
SMART: Robust and efficient fine-tuning for pre-trained natural language models through principled regularized optimization
Haoming Jiang et al · 2020
Cited alongside, same era.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al · 2020
Multi-task pre-training for plug-and-play task-oriented dialogue system
Yixuan Su et al · 2022
Later among the works it cites.
SPACE-2: Tree-structured semi-supervised contrastive pre-training for task-oriented dialog understanding
Wanwei He et al · 2022
Later among the works it cites.
Unified dialog model pre-training for task-oriented dialog understanding and generation
Wanwei He et al · 2022
Later among the works it cites.
Long time no see! open-domain conversation with long-term persona memory
Xinchao Xu et al · 2022
Later among the works it cites.
Less is more: Learning to refine dialogue history for personalized dialogue generation
Hanxun Zhong et al · 2022
Later among the works it cites.
Improving language models by retrieving from trillions of tokens
Sebastian Borgeaud et al · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Span-ConveRT: Few-shot span extraction for dialog with pretrained conversational representations
Samuel Coope et al · 2020
Cited alongside, same era.
Slot attention with value normalization for multi-domain dialogue state tracking
Yexiang Wang, Yi Guo, and Siqi Zhu · 2020
Cited alongside, same era.
MinTL: Minimalist transfer learning for task-oriented dialogue systems
Zhaojiang Lin et al · 2020
Cited alongside, same era.
Few-shot natural language generation for task-oriented dialog
Baolin Peng et al · 2020
Cited alongside, same era.
Few-shot NLG with pre-trained language model
Zhiyu Chen et al · 2020
Cited alongside, same era.
Fluent response generation for conversational question answering
Ashutosh Baheti, Alan Ritter, and Kevin Small · 2020
Cited alongside, same era.
A simple language model for task-oriented dialogue
Ehsan Hosseini-Asl et al · 2020
Cited alongside, same era.
Later among the works it cites.
Reason first, then respond: Modular generation for knowledge-infused dialogue
Leonard Adolphs et al · 2022
Later among the works it cites.
Internet-augmented dialogue generation
Mojtaba Komeili, Kurt Shuster, and Jason Weston · 2022
Later among the works it cites.
MultiWOZ 2.4: A multi-domain task-oriented dialogue dataset with essential annotation corrections to improve state tracking evaluation
Fanghua Ye, Jarana Manotumruksa, and Emine Yilmaz · 2022
Later among the works it cites.
Jamin Shin, Hangyeol Yu, Hyeongdon Moon, Andrea Madotto, and Juneyoung Park · 2022
Later among the works it cites.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models, 2023
Hugo Touvron et al · 2023
Later among the works it cites.
Chat-rec: Towards interactive and explainable llms-augmented recommender system
Yunfan Gao, Tao Sheng, Youlin Xiang, Yun Xiong, Haofen Wang, and Jiawei Zhang · 2023
Later among the works it cites.
Recent advances in deep learning based dialogue systems: A systematic survey
Jinjie Ni et al · 2023
Later among the works it cites.
End-to-end task-oriented dialogue: A survey of tasks, methods, and future directions
Libo et al. Qin · 2023
Later among the works it cites.
Gpt-4 technical report, 2023
OpenAI · 2023
Later among the works it cites.
Introducing gpts, November 2023
OpenAI · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron et al · 2023
Later among the works it cites.
Meta and microsoft introduce the next generation of llama
Meta · 2023
Later among the works it cites.
Code llama: Open foundation models for code, 2023
Baptiste Rozière et al · 2023
Later among the works it cites.
Chatglm3
THUDM · 2023
Later among the works it cites.
Gpt understands, too
Xiao Liu et al · 2023
Later among the works it cites.
Qlora: Efficient finetuning of quantized llms
Tim Dettmers et al · 2023
Later among the works it cites.
Automatic prompt augmentation and selection with chain-of-thought from labeled data
Kashun Shum, Shizhe Diao, and Tong Zhang · 2023
Later among the works it cites.
Multimodal chain-of-thought reasoning in language models
Zhuosheng Zhang et al · 2023
Later among the works it cites.
Not all languages are created equal in LLMs: Improving multilingual capability by cross-lingual-thought prompting
Haoyang Huang et al · 2023
Later among the works it cites.
Query rewriting in retrieval-augmented large language models
Xinbei Ma, Yeyun Gong, Pengcheng He, Hai Zhao, and Nan Duan · 2023
Later among the works it cites.
Enhancing retrieval-augmented large language models with iterative retrieval-generation synergy
Zhihong Shao, Yeyun Gong, Yelong Shen, Minlie Huang, Nan Duan, and Weizhu Chen · 2023
Later among the works it cites.
Zero-shot-bert-adapters: a zero-shot pipeline for unknown intent detection
Daniele Comi et al · 2023
Later among the works it cites.
Selective in-context data augmentation for intent detection using pointwise V-information
Yen-Ting Lin et al · 2023
Later among the works it cites.
CoF-CoT: Enhancing large language models with coarse-to-fine chain-of-thought prompting for multi-domain NLU tasks
Hoang Nguyen et al · 2023
Later among the works it cites.
Towards LLM-driven dialogue state tracking
Yujie Feng et al · 2023
Later among the works it cites.
Diverse retrieval-augmented in-context learning for dialogue state tracking
Brendan King and Jeffrey Flanigan · 2023
Later among the works it cites.
Prompt-based Monte-Carlo tree search for goal-oriented dialogue policy planning
Xiao Yu, Maximillian Chen, and Zhou Yu · 2023
Later among the works it cites.
Exploring zero and few-shot techniques for intent classification
Soham Parikh et al · 2023
Later among the works it cites.
Self-rag: Learning to retrieve, generate, and critique through self-reflection
Akari Asai, Zeqiu Wu, Yizhong Wang, Avirup Sil, and Hannaneh Hajishirzi · 2023
Later among the works it cites.
Learning to memorize entailment and discourse relations for persona-consistent dialogues
Ruijun Chen et al · 2023
Later among the works it cites.
Enhancing personalized dialogue generation with contrastive latent variables: Combining sparse and dense persona
Yihong Tang et al · 2023
Later among the works it cites.
Chain-of-note: Enhancing robustness in retrieval-augmented language models
Wenhao Yu, Hongming Zhang, Xiaoman Pan, Kaixin Ma, Hongwei Wang, and Dong Yu · 2023
Later among the works it cites.
MMDialog: A large-scale multi-turn dialogue dataset towards multi-modal open-domain conversation
Jiazhan Feng et al · 2023
Later among the works it cites.
From words to watts: Benchmarking the energy costs of large language model inference
Siddharth Samsi, Dan Zhao, Joseph McDonald, Baolin Li, Adam Michaleas, Michael Jones, William Bergeron, Jeremy Kepner, Devesh Tiwari, and Vijay Gadepally · 2023
Later among the works it cites.
Mind2web: Towards a generalist agent for the web
Xiang et al · 2023
Later among the works it cites.
Chain-of-exemplar: Enhancing distractor generation for multimodal educational question generation
Haohao Luo, Yang Deng, Ying Shen, See-Kiong Ng, and Tat-Seng Chua · 2024
Closest in time.
Large language model based long-tail query rewriting in taobao search
Wenjun Peng, Guiyang Li, Yue Jiang, Zilong Wang, Dan Ou, Xiaoyi Zeng, Derong Xu, Tong Xu, and Enhong Chen · 2024
Closest in time.
Fact-aware summarization with contrastive learning for few-shot dialogue state tracking
Sijie Feng, Haoxiang Su, Hongyan Xie, Di Wu, Hao Huang, and Wushour Silamu · 2024
Closest in time.
Synergizing in-context learning with hints for end-to-end task-oriented dialog systems
Vishal Vivek Saley, Rocktim Jyoti Das, Dinesh Raghu, and Mausam · 2024
Closest in time.
SynTOD: Augmented response synthesis for robust end-to-end task-oriented dialogue system
Nguyen Quang Chieu, Quang-Minh Tran, and Khac-Hoai Nam Bui · 2024
Closest in time.
Unsupervised end-to-end task-oriented dialogue with llms: The power of the noisy channel
Brendan King and Jeffrey Flanigan · 2024
Closest in time.
Leandojo: Theorem proving with retrieval-augmented language models
Kaiyu Yang, Aidan Swope, Alex Gu, Rahul Chalamala, Peiyang Song, Shixing Yu, Saad Godil, Ryan J Prenger, and Animashree Anandkumar · 2024
Closest in time.
Awq: Activation-aware weight quantization for on-device llm compression and acceleration
Ji Lin, Jiaming Tang, Haotian Tang, Shang Yang, Wei-Ming Chen, Wei-Chen Wang, Guangxuan Xiao, Xingyu Dang, Chuang Gan, and Song Han · 2024
Closest in time.
Chatdev: Communicative agents for software development
Chen Qian, Wei Liu, Hongzhang Liu, Nuo Chen, Yufan Dang, Jiahao Li, Cheng Yang, Weize Chen, Yusheng Su, Xin Cong, Juyuan Xu, Dahai Li, Zhiyuan Liu, and Maosong Sun · 2024
Closest in time.
Metagpt: Meta programming for A multi-agent collaborative framework
Sirui Hong, Mingchen Zhuge, Jonathan Chen, Xiawu Zheng, Yuheng Cheng, Jinlin Wang, Ceyao Zhang, Zili Wang, Steven Ka Shing Yau, Zijuan Lin, Liyang Zhou, Chenyu Ran, Lingfeng Xiao, Chenglin Wu, and Jürgen Schmidhuber · 2024
Closest in time.
Agentbench: Evaluating llms as agents
Xiao Liu, Hao Yu, Hanchen Zhang, Yifan Xu, Xuanyu Lei, Hanyu Lai, Yu Gu, Hangliang Ding, Kaiwen Men, Kejuan Yang, Shudan Zhang, Xiang Deng, Aohan Zeng, Zhengxiao Du, Chenhui Zhang, Sheng Shen, Tianjun Zhang, Yu Su, Huan Sun, Minlie Huang, Yuxiao Dong, and Jie Tang · 2024
Closest in time.
Lemur: Harmonizing natural language and code for language agents
Yiheng Xu, Hongjin Su, Chen Xing, Boyu Mi, Qian Liu, Weijia Shi, Binyuan Hui, Fan Zhou, Yitao Liu, Tianbao Xie, Zhoujun Cheng, Siheng Zhao, Lingpeng Kong, Bailin Wang, Caiming Xiong, and Tao Yu · 2024
Closest in time.
Prompt leakage effect and defense strategies for multi-turn llm interactions
Divyansh Agarwal, Alexander R Fabbri, Ben Risher, Philippe Laban, Shafiq Joty, and Chien-Sheng Wu · 2024
Closest in time.
Intent-driven in-context learning for few-shot dialogue state tracking
Zihao Yi, Zhe Xu, and Ying Shen · 2025
Closest in time.
Knowledge distillation on graphs: A survey
Yijun Tian, Shichao Pei, Xiangliang Zhang, Chuxu Zhang, and Nitesh V Chawla · 2025
Closest in time.
Longrope2: Near-lossless llm context window scaling
Ning Shang, Li Lyna Zhang, Siyuan Wang, Gaokai Zhang, Gilsinia Lopez, Fan Yang, Weizhu Chen, and Mao Yang · 2025
Closest in time.
Beyond single-turn: A survey on multi-turn interactions with large language models
Yubo Li, Xiaobin Shen, Xinyu Yao, Xueying Ding, Yidi Miao, Ramayya Krishnan, and Rema Padman · 2025
Closest in time.
Omnidialog: An omnipotent pre-training model for task-oriented dialogue system
Mingtao Yang, See-Kiong Ng, and Jinlan Fu · 2025
Closest in time.