Towards bidirectional hierarchical representations for attention-based neural machine translation
Baosong Yang, Derek F. Wong, Tong Xiao, Lidia S. Chao, and Jingbo Zhu · 2017
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Original
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
Top-down tree structured decoding with syntactic connections for neural machine translation and parsing
Jetic Gū, Hassan S. Shavarani, and Anoop Sarkar · 2018
Cited alongside, same era.
Scaling neural machine translation
Myle Ott, Sergey Edunov, David Grangier, and Michael Auli · 2018
Cited alongside, same era.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever · 2018
Cited alongside, same era.
Straight to the tree: Constituency parsing with neural syntactic distance
Yikang Shen, Zhouhan Lin, Athul Paul Jacob, Alessandro Sordoni, Aaron Courville, and Yoshua Bengio · 2018
Cited alongside, same era.
On tree-based neural sentence modeling
Haoyue Shi, Hao Zhou, Jiaze Chen, and Lei Li · 2018
Cited alongside, same era.
Linguistically-informed self-attention for semantic role labeling
Emma Strubell, Patrick Verga, Daniel Andor, David Weiss, and Andrew McCallum · 2018
Cited alongside, same era.
Multi-granularity self-attention for neural machine translation
Jie Hao, Xing Wang, Shuming Shi, Jinfeng Zhang, and Zhaopeng Tu · 2019
Cited alongside, same era.
über sinn und bedeutung
Gottlob Frege
Cited in the paper.