Fetching the paper…
Reading the bibliography…
Transformer-based pretrained language models (PLMs) offer unmatched performance across the majority of natural language understanding (NLU) tasks, including a body of question answering (QA) tasks.
BERT for joint intent classification and slot filling
Qian Chen, Zhu Zhuo, and Wen Wang. 2019 · 1902
Earlier work this paper cites.
RoBERTa: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
ALBERT: A lite BERT for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2020 · 1909
Earlier work this paper cites.
Distilbert, a distilled version of BERT: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Multi-domain dialogue state tracking as dynamic knowledge graph enhanced question answering
Li Zhou and Kevin Small. 2019 · 1911
Earlier work this paper cites.
Talking to machines (statistically speaking)
Steve Young. 2002 · 2002
Earlier work this paper cites.
Pre-trained models for Natural Language Processing: A survey
Xipeng Qiu, Tianxiang Sun, Yige Xu, Yunfan Shao, Ning Dai, and Xuanjing Huang. 2020 · 2003
Earlier work this paper cites.
Let’s go public! Taking a spoken dialog system to the real world
Antoine Raux, Brian Langner, Dan Bohus, Alan W Black, and Maxine Eskenazi. 2005 · 2005
Earlier work this paper cites.
Language models as few-shot learner for task-oriented dialogue systems
Andrea Madotto, Zihan Liu, Zhaojiang Lin, and Pascale Fung. 2020b · 2008
Earlier work this paper cites.
DialoGLUE: A natural language understanding benchmark for task-oriented dialogue
Shikib Mehri, Mihail Eric, and Dilek Hakkani-Tür. 2020 · 2009
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
MultiWOZ - A large-scale multi-domain wizard-of-oz dataset for task-oriented dialogue modelling
Paweł Budzianowski, Tsung-Hsien Wen, Bo-Hsiang Tseng, Iñigo Casanueva, Stefan Ultes, Osman Ramadan, and Milica Gašić. 2018 · 2018
Earlier work this paper cites.
The natural language decathlon: Multitask learning as question answering
Bryan McCann, Nitish Shirish Keskar, Caiming Xiong, and Richard Socher. 2018 · 2018
Earlier work this paper cites.
Know what you don’t know: Unanswerable questions for SQuAD
Pranav Rajpurkar, Robin Jia, and Percy Liang. 2018 · 2018
Earlier work this paper cites.
Guan-Lin Chao and Ian Lane. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
MRQA 2019 shared task: Evaluating generalization in reading comprehension
Adam Fisch, Alon Talmor, Robin Jia, Minjoon Seo, Eunsol Choi, and Danqi Chen. 2019 · 2019
Cited alongside, same era.
Dialog state tracking: A neural reading comprehension approach
Shuyang Gao, Abhishek Sethi, Sanchit Agarwal, Tagyoung Chung, and Dilek Hakkani-Tur. 2019 · 2019
Cited alongside, same era.
Parameter-efficient transfer learning for NLP
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly. 2019 · 2019
Cited alongside, same era.
Efficient intent detection with dual sentence encoders
Iñigo Casanueva, Tadas Temčinas, Daniela Gerz, Matthew Henderson, and Ivan Vulić. 2020 · 2020
Cited alongside, same era.
DIALOGPT : Large-scale generative pre-training for conversational response generation
Yizhe Zhang, Siqi Sun, Michel Galley, Yen-Chun Chen, Chris Brockett, Xiang Gao, Jianfeng Gao, Jingjing Liu, and Bill Dolan. 2020 · 2020
Later among the works it cites.
Making pre-trained language models better few-shot learners
Tianyu Gao, Adam Fisch, and Danqi Chen. 2021 · 2021
Later among the works it cites.
ConVEx: Data-efficient and few-shot slot labeling
Matthew Henderson and Ivan Vulić. 2021 · 2021
Later among the works it cites.
DS-TOD: efficient domain specialization for task oriented dialog
Chia-Chien Hung, Anne Lauscher, Simone Paolo Ponzetto, and Goran Glavaš. 2021 · 2021
Later among the works it cites.
PAQ: 65 million probably-asked questions and what you can do with them
Patrick Lewis, Yuxiang Wu, Linqing Liu, Pasquale Minervini, Heinrich Küttler, Aleksandra Piktus, Pontus Stenetorp, and Sebastian Riedel. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kevin Clark, Minh-Thang Luong, Quoc V Le, and Christopher D. Manning. 2020 · 2020
Cited alongside, same era.
Span-ConveRT: Few-shot span extraction for dialog with pretrained conversational representations
Samuel Coope, Tyler Farghly, Daniela Gerz, Ivan Vulić, and Matthew Henderson. 2020 · 2020
Cited alongside, same era.
From machine reading comprehension to dialogue state tracking: Bridging the gap
Shuyang Gao, Sanchit Agarwal, Di Jin, Tagyoung Chung, and Dilek Hakkani-Tur. 2020 · 2020
Cited alongside, same era.
TripPy: A triple copy strategy for value independent neural dialog state tracking
Michael Heck, Carel van Niekerk, Nurul Lubis, Christian Geishauser, Hsien-Chin Lin, Marco Moresi, and Milica Gasic. 2020 · 2020
Cited alongside, same era.
A simple language model for task-oriented dialogue
Ehsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz, and Richard Socher. 2020 · 2020
Cited alongside, same era.
Few-shot slot tagging with collapsed dependency transfer and label-enhanced task-adaptive projection network
Yutai Hou, Wanxiang Che, Yongkui Lai, Zhihan Zhou, Yijia Liu, Han Liu, and Ting Liu. 2020 · 2020
Cited alongside, same era.
Coach: A coarse-to-fine approach for cross-domain slot filling
Zihan Liu, Genta Indra Winata, Peng Xu, and Pascale Fung. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
Pengfei Liu, Weizhe Yuan, Jinlan Fu, Zhengbao Jiang, Hiroaki Hayashi, and Graham Neubig. 2021 · 2021
Later among the works it cites.
GenSF: Simultaneous adaptation of generative pre-trained models and slot filling
Shikib Mehri and Maxine Eskénazi. 2021 · 2021
Later among the works it cites.
Language model is all you need: Natural language understanding as question answering
Mahdi Namazifar, Alexandros Papangelis, Gokhan Tur, and Dilek Hakkani-Tür. 2021 · 2021
Later among the works it cites.
Soloist: Buildingtask bots at scale with transfer learning and machine teaching
Baolin Peng, Chunyuan Li, Jinchao Li, Shahin Shayandeh, Lars Liden, and Jianfeng Gao. 2021 · 2021
Later among the works it cites.
AdapterFusion: Non-destructive task composition for transfer learning
Jonas Pfeiffer, Aishwarya Kamath, Andreas Rückĺe, Cho Kyunghyun, and Iryna Gurevych. 2021 · 2021
Later among the works it cites.
Crossing the conversational chasm: A primer on multilingual task-oriented dialogue systems
Evgeniia Razumovskaia, Goran Glavaš, Olga Majewska, Anna Korhonen, and Ivan Vulić. 2021 · 2021
Later among the works it cites.
Recent advances in language model fine-tuning
Sebastian Ruder. 2021 · 2021
Later among the works it cites.
Task-oriented dialogue system as natural language generation
Weizhi Wang, Zhirui Zhang, Junliang Guo, Yinpei Dai, Boxing Chen, and Weihua Luo. 2021 · 2021
Later among the works it cites.
BitFit: Simple parameter-efficient fine-tuning for transformer-based masked language-models
Elad Ben Zaken, Shauli Ravfogel, and Yoav Goldberg. 2021 · 2021
Later among the works it cites.
Retrospective reader for machine reading comprehension
Zhuosheng Zhang, Junjie Yang, and Hai Zhao. 2021 · 2021
Later among the works it cites.