Tree-structured decoding with doubly-recurrent neural networks
Alvarez-Melis, D. and Jaakkola, T. S · 2017
Later among the works it cites.
Geometric deep learning: going beyond euclidean data
Bronstein, M. M., Bruna, J., LeCun, Y., Szlam, A., and Vandergheynst, P · 2017
Later among the works it cites.
A statistical approach to machine translation
Brown, P. F., Cocke, J., Pietra, S. A. D., Pietra, V. J. D., Jelinek, F., Lafferty, J. D., Mercer, R. L., and Roossin, P. S · 2017
Later among the works it cites.
Learning to parse and translate improves neural machine translation
Original
Eriguchi, A., Tsuruoka, Y., and Cho, K · 2017
Later among the works it cites.
Non-autoregressive neural machine translation
Original
Gu, J., Bradbury, J., Xiong, C., Li, V. O., and Socher, R · 2017
Later among the works it cites.
Parallel wavenet: Fast high-fidelity speech synthesis
Original
Oord, A. v. d., Li, Y., Babuschkin, I., Simonyan, K., Vinyals, O., Kavukcuoglu, K., Driessche, G. v. d., Lockhart, E., Cobo, L. C., Stimberg, F., et al · 2017
Later among the works it cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L. u., and Polosukhin, I · 2017
Later among the works it cites.
Relational inductive biases, deep learning, and graph networks
Original
Battaglia, P. W., Hamrick, J. B., Bapst, V., Sanchez-Gonzalez, A., Zambaldi, V., Malinowski, M., Tacchetti, A., Raposo, D., Santoro, A., Faulkner, R., et al · 2018
Later among the works it cites.
Language gans falling short
Original
Caccia, M., Caccia, L., Fedus, W., Larochelle, H., Pineau, J., Charlin MILA, L., and Montréal, H · 2018
Later among the works it cites.
Fast Policy Learning through Imitation and Reinforcement
Original
Cheng, C.-A., Yan, X., Wagener, N., and Boots, B · 2018
Later among the works it cites.
The importance of generation order in language modeling
Original
Ford, N., Duckworth, D., Norouzi, M., and Dahl, G. E · 2018
Later among the works it cites.
SeaRNN: Training RNNs with global-local losses
Leblond, R., Alayrac, J.-B., Osokin, A., and Lacoste-Julien, S · 2018
Later among the works it cites.
Deterministic non-autoregressive neural sequence modeling by iterative refinement
Original
Lee, J., Mansimov, E., and Cho, K · 2018
Later among the works it cites.
YiSi: A semantic machine translation evaluation metric for evaluating languages with different levels of available resources
Lo, C · 2018
Later among the works it cites.
Blockwise parallel decoding for deep autoregressive models
Stern, M., Shazeer, N., and Uszkoreit, J · 2018
Later among the works it cites.
Semi-autoregressive neural machine translation
Original
Wang, C., Zhang, J., and Chen, H · 2018
Later among the works it cites.
Loss functions for multiset prediction
Welleck, S., Yao, Z., Gai, Y., Mao, J., Zhang, Z., and Cho, K · 2018
Later among the works it cites.
Personalizing dialogue agents: I have a dog, do you have pets too?
Zhang, S., Dinan, E., Urbanek, J., Szlam, A., Kiela, D., and Weston, J · 2018
Later among the works it cites.
Texygen: A benchmarking platform for text generation models
Zhu, Y., Lu, S., Zheng, L., Guo, J., Zhang, W., Wang, J., and Yu, Y · 2018
Later among the works it cites.
Novel positional encodings to enable tree-structured transformers
Shiv, V. L. and Quirk, C · 2019
Closest in time.