Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Original
Diederik P. Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Cited alongside, same era.
Structured prediction energy networks
David Belanger and Andrew McCallum. 2016 · 2016
Cited alongside, same era.
Neural architectures for named entity recognition
Guillaume Lample, Miguel Ballesteros, Sandeep Subramanian, Kazuya Kawakami, and Chris Dyer. 2016 · 2016
Cited alongside, same era.
End-to-end sequence labeling via bi-directional LSTM-CNNs-CRF
Xuezhe Ma and Eduard Hovy. 2016 · 2016
Cited alongside, same era.
Improved techniques for training GANs
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, and Xi Chen. 2016 · 2016
Cited alongside, same era.
Wasserstein generative adversarial networks
Martín Arjovsky, Soumith Chintala, and Léon Bottou. 2017 · 2017
Cited alongside, same era.
End-to-end learning for structured prediction energy networks
David Belanger, Bishan Yang, and Andrew McCallum. 2017 · 2017
Cited alongside, same era.
Improved training of Wasserstein GANs
Ishaan Gulrajani, Faruk Ahmed, Martin Arjovsky, Vincent Dumoulin, and Aaron C Courville. 2017 · 2017
Cited alongside, same era.
Learning to embed words in context for syntactic tasks
Lifu Tu, Kevin Gimpel, and Karen Livescu. 2017 · 2017
Cited alongside, same era.
ENGINE: Energy-based inference networks for non-autoregressive machine translation
Lifu Tu, Richard Yuanzhe Pang, Sam Wiseman, and Kevin Gimpel. 2020b
Cited in the paper.