Fetching the paper…
Reading the bibliography…
While current dialogue systems like ChatGPT have made significant advancements in text-based interactions, they often overlook the potential of other modalities in enhancing the overall user experience.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 1901
Earlier work this paper cites.
An architecture for dialogue management, context tracking, and pragmatic adaptation in spoken dialogue systems
Susann LuperFoy, Dan Loehr, David Duff, Keith Miller, Florence Reeder, and Lisa Harper. 1998 · 1998
Earlier work this paper cites.
The OpenCV Library
G. Bradski. 2000 · 2000
Earlier work this paper cites.
Human-robot interaction based on spoken natural language dialogue
Dimitris Spiliotopoulos, Ion Androutsopoulos, and Costantine Spyropoulos. 2001 · 2001
Earlier work this paper cites.
Multimodal dialogue systems
Alexander I Rudnicky. 2005 · 2005
Earlier work this paper cites.
Review of spoken dialogue systems
Ramón López-Cózar, Zoraida Callejas, David Griol, and Jose Quesada. 2015 · 2015
Earlier work this paper cites.
Building end-to-end dialogue systems using generative hierarchical neural network models
Iulian Serban, Alessandro Sordoni, Yoshua Bengio, Aaron C. Courville, and Joelle Pineau. 2015 · 2015
Earlier work this paper cites.
Ticktock: A non-goal-oriented multimodal dialog system with engagement awareness
Zhou Yu, Alexandros Papangelis, and Alexander I. Rudnicky. 2015 · 2015
Earlier work this paper cites.
Tensorflow: A system for large-scale machine learning
Martín Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving, Michael Isard, et al. 2016 · 2016
Earlier work this paper cites.
SSD: single shot multibox detector
Wei Liu, Dragomir Anguelov, Dumitru Erhan, Christian Szegedy, Scott E. Reed, Cheng-Yang Fu, and Alexander C. Berg. 2016 · 2016
Earlier work this paper cites.
A survey on dialogue systems: Recent advances and new frontiers
Hongshen Chen, Xiaorui Liu, Dawei Yin, and Jiliang Tang. 2017 · 2017
Earlier work this paper cites.
Audio-driven facial animation by joint end-to-end learning of pose and emotion
Tero Karras, Timo Aila, Samuli Laine, Antti Herva, and Jaakko Lehtinen. 2017 · 2017
Cited alongside, same era.
Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit
Christophe Veaux, Junichi Yamagishi, and Kirsten MacDonald. 2017 · 2017
Cited alongside, same era.
Multimodal HALEF: An Open-Source Modular Web-Based Multimodal Dialog Framework , pages 233–244
Zhou Yu, Vikram Ramanarayanan, Robert Mundkowsky, Patrick Lange, Alexei Ivanov, Alan Black, and David Suendermann-Oeft. 2017 · 2017
Cited alongside, same era.
Flask web development: developing web applications with python
Miguel Grinberg. 2018 · 2018
Cited alongside, same era.
Dialog systems and chatbots
Daniel Jurafsky and James H Martin. 2018 · 2018
Cited alongside, same era.
Sentiment adaptive end-to-end dialog systems
Weiyan Shi and Zhou Yu. 2018 · 2018
ChatGPT: A large-scale pretrained conversation model
OpenAI. 2021 · 2021
Later among the works it cites.
Hyperextended lightface: A facial attribute analysis framework
Sefik Ilkin Serengil and Alper Ozpinar. 2021 · 2021
Later among the works it cites.
Silero vad: pre-trained enterprise-grade voice activity detector (vad), number detector and language classifier
Silero Team. 2021 · 2021
Later among the works it cites.
Towards building a spoken dialogue system for argument exploration
Annalena Aicher, Nadine Gerstenlauer, Isabel Feustel, Wolfgang Minker, and Stefan Ultes. 2022 · 2022
Later among the works it cites.
Emosen: Generating sentiment and emotion controlled responses in a multimodal dialogue system
Mauajama Firdaus, Hardik Chauhan, Asif Ekbal, and Pushpak Bhattacharyya. 2022 · 2022
Later among the works it cites.
Kurt: A household assistance robot capable of proactive dialogue
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Personalizing dialogue agents: I have a dog, do you have pets too?
Saizheng Zhang, Emily Dinan, Jack Urbanek, Arthur Szlam, Douwe Kiela, and Jason Weston. 2018 · 2018
Cited alongside, same era.
Multi-step reasoning via recurrent dual attention for visual dialog
Zhe Gan, Yu Cheng, Ahmed Kholy, Linjie Li, Jingjing Liu, and Jianfeng Gao. 2019 · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019 · 2019
Cited alongside, same era.
Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech
Jaehyeon Kim, Jungil Kong, and Juhee Son. 2021 · 2021
Cited alongside, same era.
Matthias Kraus, Nicolas Wagner, Wolfgang Minker, Ankita Agrawal, Artur Schmidt, Pranav Krishna Prasad, and Wolfgang Ertel. 2022 · 2022
Later among the works it cites.
Transformer-based multimodal infusion dialogue systems
Bo Liu, Lejian He, Yafei Liu, Tianyao Yu, Yuejia Xiang, Li Zhu, and Weijian Ruan. 2022 · 2022
Later among the works it cites.
Data augmentation with paraphrase generation and entity extraction for multimodal dialogue system
Eda Okur, Saurav Sahay, and Lama Nachman. 2022 · 2022
Later among the works it cites.
Robust speech recognition via large-scale weak supervision
Alec Radford, Jong Wook Kim, Tao Xu, Greg Brockman, Christine McLeavey, and Ilya Sutskever. 2022 · 2022
Later among the works it cites.
Neural codec language models are zero-shot text to speech synthesizers
Chengyi Wang, Sanyuan Chen, Yu Wu, Ziqiang Zhang, Long Zhou, Shujie Liu, Zhuo Chen, Yanqing Liu, Huaming Wang, Jinyu Li, Lei He, Sheng Zhao, and Furu Wei. 2023 · 2023
Closest in time.