Fetching the paper…
Reading the bibliography…
In this paper, we propose Dynamic Self-Attention (DSA), a new self-attention mechanism for sentence embedding.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey E. Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart van Merrienboer, Çaglar Gülçehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D. Manning. 2014 · 2014
Earlier work this paper cites.
Learning natural language inference using bidirectional LSTM model and inner-attention
Yang Liu, Chengjie Sun, Lei Lin, and Xiaolong Wang. 2016 · 2016
Earlier work this paper cites.
Natural language inference by tree-based convolution and heuristic matching
Lili Mou, Rui Men, Ge Li, Yan Xu, Lu Zhang, Rui Yan, and Zhi Jin. 2016 · 2016
Earlier work this paper cites.
Attention-based convolutional neural network for semantic relation extraction
Yatian Shen and Xuanjing Huang. 2016 · 2016
Cited alongside, same era.
Hierarchical attention networks for document classification
Zichao Yang, Diyi Yang, Chris Dyer, Xiaodong He, Alexander J. Smola, and Eduard H. Hovy. 2016 · 2016
Cited alongside, same era.
Supervised learning of universal sentence representations from natural language inference data
Alexis Conneau, Douwe Kiela, Holger Schwenk, Loïc Barrault, and Antoine Bordes. 2017 · 2017
Cited alongside, same era.
Convolutional sequence to sequence learning
Jonas Gehring, Michael Auli, David Grangier, Denis Yarats, and Yann N. Dauphin. 2017 · 2017
Cited alongside, same era.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Laurens van der Maaten, and Kilian Q. Weinberger. 2017 · 2017
Cited alongside, same era.
Distance-based self-attention network for natural language inference
A structured self-attentive sentence embedding
Zhouhan Lin, Minwei Feng, Cícero Nogueira dos Santos, Mo Yu, Bing Xiang, Bowen Zhou, and Yoshua Bengio. 2017 · 2017
Later among the works it cites.
Shortcut-stacked sentence encoders for multi-domain inference
Yixin Nie and Mohit Bansal. 2017 · 2017
Later among the works it cites.
Dynamic routing between capsules
Sara Sabour, Nicholas Frosst, and Geoffrey E. Hinton. 2017 · 2017
Later among the works it cites.
Learning to compose task-specific tree structures
Jihun Choi, Kang Min Yoo, and Sang-goo Lee. 2018 · 2018
Closest in time.
Investigating capsule networks with dynamic routing for text classification
Wei Zhao, Jianbo Ye, Min Yang, Zeyang Lei, Soufei Zhang, and Zhou Zhao. 2018 · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jinbae Im and Sungzoon Cho. 2017 · 2017
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio
Cited in the paper.
Disan: Directional self-attention network for rnn/cnn-free language understanding
Tao Shen, Tianyi Zhou, Guodong Long, Jing Jiang, Shirui Pan, and Chengqi Zhang. 2018a
Cited in the paper.
Reinforced self-attention network: A hybrid of hard and soft attention for sequence modeling
Tao Shen, Tianyi Zhou, Guodong Long, Jing Jiang, Sen Wang, and Chengqi Zhang. 2018b
Cited in the paper.