Fetching the paper…
Reading the bibliography…
Self-attention mechanisms have achieved great success on a variety of NLP tasks due to its flexibility of capturing dependency between arbitrary positions in a sequence.
Automatic evaluation of machine translation quality using longest common subsequence and skip-bigram statistics
Chin-Yew Lin and Franz Josef Och. 2004 · 2004
Earlier work this paper cites.
Query based summarization using non-negative matrix factorization
Sun Park, Ju-Hong Lee, Chan-Min Ahn, Jun Sik Hong, and Seok-Ju Chun. 2006 · 2006
Earlier work this paper cites.
Fastsum: fast and accurate query-based multi-document summarization
Frank Schilder and Ravikumar Kondadadi. 2008 · 2008
Earlier work this paper cites.
Multi-document summarization using cluster-based link analysis
Xiaojun Wan and Jianwu Yang. 2008 · 2008
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Earlier work this paper cites.
Using query expansion in graph-based approach for query-focused multi-document summarization
Lin Zhao, Lide Wu, and Xuanjing Huang. 2009 · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
A neural attention model for abstractive sentence summarization
Alexander M Rush, Sumit Chopra, and Jason Weston. 2015 · 2015
Earlier work this paper cites.
Neural responding machine for short-text conversation
Lifeng Shang, Zhengdong Lu, and Hang Li. 2015 · 2015
Earlier work this paper cites.
End-to-end memory networks
Sainbayar Sukhbaatar, Jason Weston, Rob Fergus, et al. 2015 · 2015
Cited alongside, same era.
Attsum: Joint learning of focusing and summarization with neural attention
Ziqiang Cao, Wenjie Li, Sujian Li, Furu Wei, and Yanran Li. 2016 · 2016
Cited alongside, same era.
Learning natural language inference using bidirectional lstm model and inner-attention
Yang Liu, Chengjie Sun, Lei Lin, and Xiaolong Wang. 2016 · 2016
Cited alongside, same era.
Query-based abstractive summarization using neural networks
Johan Hasselqvist, Niklas Helmertz, and Mikael Kågebäck. 2017 · 2017
Cited alongside, same era.
Reinforced mnemonic reader for machine comprehension
Minghao Hu, Yuxing Peng, and Xipeng Qiu. 2017 · 2017
Cited alongside, same era.
Abstractive document summarization with a graph-based attentional neural model
Jiwei Tan, Xiaojun Wan, and Jianguo Xiao. 2017 · 2017
Later among the works it cites.
Attention is all you need
Ashish Vaswani, Shazeer, Noam, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Tal Baumel, Matan Eyal, and Michael Elhadad. 2018 · 2018
Later among the works it cites.
Generating wikipedia by summarizing long sequences
Peter J. Liu, Mohammad Saleh, Etienne Pot, Ben Goodrich, Ryan Sepassi, Lukasz Kaiser, and Noam Shazeer. 2018 · 2018
Later among the works it cites.
Bi-directional block self-attention for fast and memory-efficient sequence modeling
Tao Shen, Tianyi Zhou, Guodong Long, Jing Jiang, and Chengqi Zhang. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A structured self-attentive sentence embedding
Zhouhan Lin, Minwei Feng, Cicero Nogueira dos Santos, Mo Yu, Bing Xiang, Bowen Zhou, and Yoshua Bengio. 2017 · 2017
Cited alongside, same era.
Diversity driven attention model for query-based abstractive summarization
Preksha Nema, Mitesh Khapra, Anirban Laha, and Balaraman Ravindran. 2017 · 2017
Cited alongside, same era.
Get to the point: Summarization with pointer-generator networks
Abigail See, Peter J. Liu, and Christopher D. Manning. 2017 · 2017
Cited alongside, same era.
Disan: Directional self-attention network for rnn/cnn-free language understanding
Tao Shen, Tianyi Zhou, Guodong Long, Jing Jiang, Shirui Pan, and Chengqi Zhang. 2017 · 2017
Cited alongside, same era.
Enhancing sentence embedding with generalized pooling
Qian Chen, Zhen-Hua Ling, and Xiaodan Zhu. 2018a
Cited in the paper.
Neural natural language inference models enhanced with external knowledge
Qian Chen, Xiaodan Zhu, Zhen-Hua Ling, Diana Inkpen, and Si Wei. 2018b
Cited in the paper.
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W Cohen, Ruslan Salakhutdinov, and Christopher D Manning. 2018 · 2018
Later among the works it cites.
Universal transformers
Mostafa Dehghani, Stephan Gouws, Oriol Vinyals, Jakob Uszkoreit, and Lukasz Kaiser. 2019 · 2019
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.