Fetching the paper…
Reading the bibliography…
We develop a chatbot using Deep Bidirectional Transformer models (BERT) to handle client questions in financial investment customer service.
Eliza—a computer program for the study of natural language communication between man and machine
Weizenbaum, J · 1966
Earlier work this paper cites.
Artificial paranoia
Colby, K. M., Weber, S., and Hilf, F. D · 1971
Earlier work this paper cites.
Finding structure in time
Elman, J. L · 1990
Earlier work this paper cites.
A practical bayesian framework for backpropagation networks
MacKay, D · 1992
Earlier work this paper cites.
A new algorithm for data compression
Gage, P · 1994
Earlier work this paper cites.
Bayesian learning for neural networks
Neal, R · 1995
Earlier work this paper cites.
A unified architecture for natural language processing: Deep neural networks with multitask learning
Collobert, R. and Weston, J · 2008
Earlier work this paper cites.
Recurrent neural network based language model
Mikolov, T., Karafiát, M., Burget, L., Cernocký, J., and Khudanpur, S · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
Collobert, R., Weston, J., Bottou, L., Karlen, M., Kavukcuoglu, K., and Kuksa, P. P · 2011
Earlier work this paper cites.
Practical variational inference for neural networks
Graves, A · 2011
Earlier work this paper cites.
Bayesian learning via stochastic gradient langevin dynamics
Welling, M. and Teh, Y. W · 2011
Earlier work this paper cites.
Syntactic annotations for the google books ngram corpus
Lin, Y., Michel, J.-B., Aiden, E. L., Orwant, J., Brockman, W., and Petrov, S · 2012
Earlier work this paper cites.
Japanese and korean voice search
Schuster, M. and Nakajima, K · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Mikolov, T., Sutskever, I., Chen, K., Corrado, G., and Dean, J · 2013
Earlier work this paper cites.
Deep convolutional neural networks for sentiment analysis of short texts
dos Santos, C. and Gatti, M · 2014
Earlier work this paper cites.
Learning character-level representations for part-of-speech tagging
Dos Santos, C. N. and Zadrozny, B · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y · 2015
Earlier work this paper cites.
Weight uncertainty in neural network
Blundell, C., Cornebise, J., Kavukcuoglu, K., and Wierstra, D · 2015
Earlier work this paper cites.
Semi-supervised sequence learning
Dai, A. M. and Le, Q. V · 2015
Earlier work this paper cites.
Probabilistic backpropagation for scalable learning of bayesian neural networks
Hernández-Lobato, J. M. and Adams, R. P · 2015
Cited alongside, same era.
Character-aware neural language models
Kim, Y., Jernite, Y., Sontag, D. A., and Rush, A. M · 2015
Cited alongside, same era.
Bayesian reinforcement learning
et al., M. G · 2016
Cited alongside, same era.
Dropout as bayesian approximation: representing model uncertainty in deep learning
Gal, Y. and Ghahramani, Z · 2016
Cited alongside, same era.
Threshold learning for optimal decision making
Lepora, N. F · 2016
Cited alongside, same era.
Learning weight uncertainty with stochastic gradient mcmc for shape classification
Li, C., Stevens, A., Chen, C., Pu, Y., Gan, Z., and Carin, L · 2016
Cited alongside, same era.
Sampling-based bayesian inference with gradient uncertainty
Park, C., Kim, J., Ha, S. H., and Lee, J · 2018
Later among the works it cites.
Uncertainty in neural networks: Bayesian ensembling
Pearce, T., Zaki, M., Brintrup, A., and Neely, A · 2018
Later among the works it cites.
Deep contextualized word representations
Peters, M., Neumann, M., Iyyer, M., Gardner, M., Clark, C., Lee, K., and Zettlemoyer, L · 2018
Later among the works it cites.
Improving language understanding by generative pre-training
Radford, A. and Sutskever, I · 2018
Later among the works it cites.
Deep learning for self-driving cars: Chances and challenges
Rao, Q. and Frtunikj, J · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural machine translation of rare words with subword units
Sennrich, R., Haddow, B., and Birch, A · 2016
Cited alongside, same era.
Deep-text-corrector
Atpaino · 2017
Cited alongside, same era.
Simple and scalable predictive uncertainty estimation using deep ensembles
B. Lakshminarayanan, A. Prtizel, C. B · 2017
Cited alongside, same era.
Rasa: Open source language understanding and dialogue management
Bocklisch, T., Faulkner, J., Pawlowski, N., and Nichol, A · 2017
Cited alongside, same era.
On calibration of modern neural networks
et al., C. G · 2017
Cited alongside, same era.
Avoiding echo-responses in a retrieval-based conversation system
Fedorenko, D. G., Smetanin, N., and Rodichev, A · 2017
Cited alongside, same era.
V. Kuleshov, N. F. and Ermon, S · 2018
Later among the works it cites.
The design and implementation of xiaoice, an empathetic social chatbot
Zhou, L., Gao, J., Li, D., and Shum, H · 2018
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2019
Later among the works it cites.
Selectivenet: A deep neural network with an integrated reject option
Geifman, Y · 2019
Later among the works it cites.
Cross-lingual language model pretraining
Lample, G. and Conneau, A · 2019
Later among the works it cites.
An evaluation dataset for intent classification and out-of-scope prediction
Larson, S., Mahendran, A., Peper, J. J., Clarke, C., Lee, A., Hill, P., Kummerfeld, J. K., Leach, K., Laurenzano, M. A., Tang, L., and Mars, J · 2019
Later among the works it cites.
Roberta: A robustly optimized BERT pretraining approach
Liu, Y., Ott, M., Goyal, N., Du, J., Joshi, M., Chen, D., Levy, O., Lewis, M., Zettlemoyer, L., and Stoyanov, V · 2019
Later among the works it cites.
A simple baseline for bayesian uncertainty in deep learning
Maddox, W., Garipov, T., Izmailov, P., Vetrov, D. P., and Wilson, A. G · 2019
Later among the works it cites.
Towards calibrated and scalable uncertainty representations for neural networks
Seedat, N. and Kanan, C · 2019
Later among the works it cites.
The state of chatbots in 2019
Sethi, S · 2019
Later among the works it cites.
A comprehensive guide to bayesian convolutional neural network with variational inference
Shridhar, K., Laumann, F., and Liwicki, M · 2019
Later among the works it cites.
Distilling task-specific knowledge from BERT into simple neural networks
Tang, R., Lu, Y., Liu, L., Mou, L., Vechtomova, O., and Lin, J · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Yang, Z., Dai, Z., Yang, Y., Carbonell, J. G., Salakhutdinov, R., and Le, Q. V · 2019
Later among the works it cites.
Towards a human-like open-domain chatbot, 2020
Adiwardana, D., Luong, M.-T., So, D. R., Hall, J., Fiedel, N., Thoppilan, R., Yang, Z., Kulshreshtha, A., Nemade, G., Lu, Y., and Le, Q. V · 2020
Closest in time.