Fetching the paper…
Reading the bibliography…
Many natural language processing tasks solely rely on sparse dependencies between a few tokens in a sentence.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Hybrid speech recognition with deep bidirectional lstm
Alex Graves, Navdeep Jaitly, and Abdel-rahman Mohamed · 2013
Earlier work this paper cites.
The meaning factory: Formal semantics for recognizing textual entailment and determining semantic similarity
Johannes Bjerva, Johan Bos, Rob Van der Goot, and Malvina Nissim · 2014
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim · 2014
Earlier work this paper cites.
A sick cure for the evaluation of compositional distributional semantic models
Marco Marelli, Stefano Menini, Marco Baroni, Luisa Bentivogli, Raffaella Bernardi, and Roberto Zamparelli · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D. Manning · 2014
Earlier work this paper cites.
Grounded compositional semantics for finding and describing images with sentences
Richard Socher, Andrej Karpathy, Quoc V Le, Christopher D Manning, and Andrew Y Ng · 2014
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey E Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Earlier work this paper cites.
Ecnu: One stone two birds: Ensemble of heterogenous measures for semantic relatedness and textual entailment
Jiang Zhao, Tiantian Zhu, and Man Lan · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning · 2015
Earlier work this paper cites.
Effective approaches to attention-based neural machine translation
Minh-Thang Luong, Hieu Pham, and Christopher D Manning · 2015
Earlier work this paper cites.
Neural responding machine for short-text conversation
Lifeng Shang, Zhengdong Lu, and Hang Li · 2015
Cited alongside, same era.
Rupesh Kumar Srivastava, Klaus Greff, and Jürgen Schmidhuber · 2015
Cited alongside, same era.
Improved semantic representations from tree-structured long short-term memory networks
Kai Sheng Tai, Richard Socher, and Christopher D Manning · 2015
Cited alongside, same era.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio · 2015
Cited alongside, same era.
A fast unified model for parsing and sentence understanding
Samuel R Bowman, Jon Gauthier, Abhinav Rastogi, Raghav Gupta, Christopher D Manning, and Christopher Potts · 2016
Cited alongside, same era.
More is less: A more complicated network with less inference complexity
Xuanyi Dong, Junshi Huang, Yi Yang, and Shuicheng Yan · 2017
Later among the works it cites.
Convolutional sequence to sequence learning
Jonas Gehring, Michael Auli, David Grangier, Denis Yarats, and Yann N Dauphin · 2017
Later among the works it cites.
Reinforced mnemonic reader for machine comprehension
Minghao Hu, Yuxing Peng, and Xipeng Qiu · 2017
Later among the works it cites.
End-to-end task-completion neural dialogue systems
Xuijun Li, Yun-Nung Chen, Lihong Li, and Jianfeng Gao · 2017
Later among the works it cites.
End-to-end adversarial memory network for cross-domain sentiment classification
Zheng Li, Yu Zhang, Ying Wei, Yuxiang Wu, and Qiang Yang · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yann N Dauphin, Angela Fan, Michael Auli, and David Grangier · 2016
Cited alongside, same era.
Dual learning for machine translation
Di He, Yingce Xia, Tao Qin, Liwei Wang, Nenghai Yu, Tieyan Liu, and Wei-Ying Ma · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Rationalizing neural predictions
Tao Lei, Regina Barzilay, and Tommi Jaakkola · 2016
Cited alongside, same era.
Learning natural language inference using bidirectional lstm model and inner-attention
Yang Liu, Chengjie Sun, Lei Lin, and Xiaolong Wang · 2016
Cited alongside, same era.
Learning to compose words into sentences with reinforcement learning
Dani Yogatama, Phil Blunsom, Chris Dyer, Edward Grefenstette, and Wang Ling · 2016
Cited alongside, same era.
Recurrent neural network-based sentence encoder with gated attention for natural language inference
Qian Chen, Xiaodan Zhu, Zhen-Hua Ling, Si Wei, Hui Jiang, and Diana Inkpen · 2017
Cited alongside, same era.
Later among the works it cites.
Neural semantic encoders
Tsendsuren Munkhdalai and Hong Yu · 2017
Later among the works it cites.
Shortcut-stacked sentence encoders for multi-domain inference
Yixin Nie and Mohit Bansal · 2017
Later among the works it cites.
Bidirectional attention flow for machine comprehension
Minjoon Seo, Aniruddha Kembhavi, Ali Farhadi, and Hannaneh Hajishirzi · 2017
Later among the works it cites.
Reasonet: Learning to stop reading in machine comprehension
Yelong Shen, Po-Sen Huang, Jianfeng Gao, and Weizhu Chen · 2017
Later among the works it cites.
Attention is all you need
Ashish Vaswani, Shazeer, Noam, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Later among the works it cites.
Adams Wei Yu, Hongrae Lee, and Quoc V Le · 2017
Later among the works it cites.
Sentence simplification with deep reinforcement learning
Xingxing Zhang and Mirella Lapata · 2017
Later among the works it cites.
Semantic structure-based word embedding by incorporating concept convergence and word divergence
Qian Liu, Heyan Huang, Guangquan Zhang, Yang Gao, Junyu Xuan, and Jie Lu · 2018
Closest in time.
Disan: Directional self-attention network for rnn/cnn-free language understanding
Tao Shen, Tianyi Zhou, Guodong Long, Jing Jiang, Shirui Pan, and Chengqi Zhang · 2018
Closest in time.