Fetching the paper…
Reading the bibliography…
Sequence to sequence learning has recently emerged as a new paradigm in supervised learning.
Building a large annotated corpus of english: The penn treebank
Marcus, Mitchell P., Marcinkiewicz, Mary Ann, and Santorini, Beatrice · 1993
Earlier work this paper cites.
Is learning the n-th thing any easier than learning the first?
Thrun, Sebastian · 1996
Earlier work this paper cites.
Multitask learning
Caruana, Rich · 1997
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, Kishore, Roukos, Salim, Ward, Todd, and jing Zhu, Wei · 2002
Earlier work this paper cites.
Regularized multi–task learning
Evgeniou, Theodoros and Pontil, Massimiliano · 2004
Earlier work this paper cites.
A framework for learning predictive structures from multiple tasks and unlabeled data
Ando, Rie Kubota and Zhang, Tong · 2005
Earlier work this paper cites.
Multi-task feature learning
Argyriou, Andreas, Evgeniou, Theodoros, and Pontil, Massimiliano · 2007
Earlier work this paper cites.
Learning task grouping and overlap in multi-task learning
Kumar, Abhishek and III, Hal Daumé · 2012
Earlier work this paper cites.
Multilingual acoustic models using distributed deep neural networks
Heigold, Georg, Vanhoucke, Vincent, Senior, Alan, Nguyen, Patrick, Ranzato, Marc’Aurelio, Devin, Matthieu, and Dean, Jeffrey · 2013
Cited alongside, same era.
Cross-language knowledge transfer using multilingual deep neural network with shared hidden layers
Huang, Jui-Ting, Li, Jinyu, Yu, Dong, Deng, Li, and Gong, Yifan · 2013
Cited alongside, same era.
Recurrent continuous translation models
Kalchbrenner, Nal and Blunsom, Phil · 2013
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Cho, Kyunghyun, van Merrienboer, Bart, Gulcehre, Caglar, Bougares, Fethi, Schwenk, Holger, and Bengio, Yoshua · 2014
Cited alongside, same era.
DeCAF: A deep convolutional activation feature for generic visual recognition, 2014
Donahue, Jeff, Jia, Yangqing, Vinyals, Oriol, Hoffman, Judy, Zhang, Ning, Tzeng, Eric, and Darrell, Trevor · 2014
Cited alongside, same era.
Semi-supervised sequence learning
Dai, Andrew M. and Le, Quoc V · 2015
Closest in time.
Multi-task learning for multiple language translation
Dong, Daxiang, Wu, Hua, He, Wei, Yu, Dianhai, and Wang, Haifeng · 2015
Closest in time.
On using monolingual corpora in neural machine translation
Gulcehre, Caglar, Firat, Orhan, Xu, Kelvin, Cho, Kyunghyun, Barrault, Loic, Lin, Huei-Chi, Bougares, Fethi, Schwenk, Holger, and Bengio, Yoshua · 2015
Closest in time.
Skip-thought vectors
Kiros, Ryan, Zhu, Yukun, Salakhutdinov, Ruslan, Zemel, Richard S., Torralba, Antonio, Urtasun, Raquel, and Fidler, Sanja · 2015
Closest in time.
Representation learning using multi-task deep neural networks for semantic classification and information retrieval
Liu, Xiaodong, Gao, Jianfeng, He, Xiaodong, Deng, Li, Duh, Kevin, and Wang, Ye-Yi · 2015
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dropout improves recurrent neural networks for handwriting recognition
Pham, Vu, Bluche, Théodore, Kermorvant, Christopher, and Louradour, Jérôme · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc V · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2015
Cited alongside, same era.
Findings of the 2015 workshop on statistical machine translation
Bojar, Ondřej, Chatterjee, Rajen, Federmann, Christian, Haddow, Barry, Huck, Matthias, Hokamp, Chris, Koehn, Philipp, Logacheva, Varvara, Monz, Christof, Negri, Matteo, Post, Matt, Scarton, Carolina, Specia, Lucia, and Turchi, Marco · 2015
Cited alongside, same era.
On using very large target vocabulary for neural machine translation
Jean, Sébastien, Cho, Kyunghyun, Memisevic, Roland, and Bengio, Yoshua
Cited in the paper.
Montreal neural machine translation systems for WMT’15
Jean, Sébastien, Firat, Orhan, Cho, Kyunghyun, Memisevic, Roland, and Bengio, Yoshua
Cited in the paper.
Effective approaches to attention-based neural machine translation
Luong, Minh-Thang, Pham, Hieu, and Manning, Christopher D
Cited in the paper.
Luong, Minh-Thang and Manning, Christopher D · 2015
Closest in time.
Show, attend and tell: Neural image caption generation with visual attention
Xu, Kelvin, Ba, Jimmy, Kiros, Ryan, Cho, Kyunghyun, Courville, Aaron C., Salakhutdinov, Ruslan, Zemel, Richard S., and Bengio, Yoshua · 2015
Closest in time.
Exploring the limits of language modeling
Jozefowicz, R., Vinyals, O., Schuster, M., Shazeer, N., and Wu, Y · 2016
Closest in time.