Fetching the paper…
Reading the bibliography…
Many state-of-the-art neural models for NLP are heavily parameterized and thus memory inefficient.
Quaternion knowledge graph embeddings
Shuai Zhang, Yi Tay, Lina Yao, and Qi Liu. 2019a · 1904
Earlier work this paper cites.
Quaternion collaborative filtering for recommendation
Shuai Zhang, Lina Yao, Lucas Vinh Tran, Aston Zhang, and Yi Tay. 2019b · 1906
Earlier work this paper cites.
Holographic reduced representations: Convolution algebra for compositional distributed representations
Tony Plate. 1991 · 1991
Earlier work this paper cites.
Learning word vectors for sentiment analysis
Andrew L. Maas, Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. 2011 · 2011
Earlier work this paper cites.
Improving word representations via global context and multiple word prototypes
Eric H Huang, Richard Socher, Christopher D Manning, and Andrew Y Ng. 2012 · 2012
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D Manning, Andrew Ng, and Christopher Potts. 2013 · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning. 2015 · 2015
Earlier work this paper cites.
Do multi-sense embeddings improve natural language understanding?
Jiwei Li and Dan Jurafsky. 2015 · 2015
Earlier work this paper cites.
The ubuntu dialogue corpus: A large dataset for research in unstructured multi-turn dialogue systems
Ryan Lowe, Nissan Pow, Iulian Serban, and Joelle Pineau. 2015 · 2015
Earlier work this paper cites.
Efficient non-parametric estimation of multiple embeddings per word in vector space
Arvind Neelakantan, Jeevan Shankar, Alexandre Passos, and Andrew McCallum. 2015 · 2015
Earlier work this paper cites.
Tensorflow: A system for large-scale machine learning
Martín Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving, Michael Isard, et al. 2016 · 2016
Earlier work this paper cites.
Unitary evolution recurrent neural networks
Martin Arjovsky, Amar Shah, and Yoshua Bengio. 2016 · 2016
Earlier work this paper cites.
Enhanced lstm for natural language inference
Qian Chen, Xiaodan Zhu, Zhenhua Ling, Si Wei, Hui Jiang, and Diana Inkpen. 2016 · 2016
Cited alongside, same era.
Associative long short-term memory
Ivo Danihelka, Greg Wayne, Benigno Uria, Nal Kalchbrenner, and Alex Graves. 2016 · 2016
Cited alongside, same era.
Assessing the ability of lstms to learn syntax-sensitive dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg. 2016 · 2016
Cited alongside, same era.
Holographic embeddings of knowledge graphs
Maximilian Nickel, Lorenzo Rosasco, and Tomaso Poggio. 2016 · 2016
Cited alongside, same era.
A decomposable attention model for natural language inference
Ankur P Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit. 2016 · 2016
Cited alongside, same era.
Quaternion convolutional neural networks for detection and localization of 3d sound events
Danilo Comminiello, Marco Lella, Simone Scardapane, and Aurelio Uncini. 2018 · 2018
Later among the works it cites.
Mostafa Dehghani, Stephan Gouws, Oriol Vinyals, Jakob Uszkoreit, and Łukasz Kaiser. 2018 · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Later among the works it cites.
Scitail: A textual entailment dataset from science question answering
Tushar Khot, Ashish Sabharwal, and Peter Clark. 2018 · 2018
Later among the works it cites.
Quaternet: A quaternion-based recurrent model for human motion
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bidirectional attention flow for machine comprehension
Minjoon Seo, Aniruddha Kembhavi, Ali Farhadi, and Hannaneh Hajishirzi. 2016 · 2016
Cited alongside, same era.
Frustratingly short attention spans in neural language modeling
Michał Daniluk, Tim Rocktäschel, Johannes Welbl, and Sebastian Riedel. 2017 · 2017
Cited alongside, same era.
Chase Gaudet and Anthony Maida. 2017 · 2017
Cited alongside, same era.
A continuously growing dataset of sentential paraphrases
Wuwei Lan, Siyu Qiu, Hua He, and Wei Xu. 2017 · 2017
Cited alongside, same era.
Chiheb Trabelsi, Olexa Bilaniuk, Ying Zhang, Dmitriy Serdyuk, Sandeep Subramanian, João Felipe Santos, Soroush Mehri, Negar Rostamzadeh, Yoshua Bengio, and Christopher J Pal. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Bilateral multi-perspective matching for natural language sentences
Zhiguo Wang, Wael Hamza, and Radu Florian. 2017 · 2017
Cited alongside, same era.
Dario Pavllo, David Grangier, and Michael Auli. 2018 · 2018
Later among the works it cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever. 2018 · 2018
Later among the works it cites.
Hermitian co-attention networks for text matching in asymmetrical domains
Yi Tay, Anh Tuan Luu, and Siu Cheung Hui. 2018 · 2018
Later among the works it cites.
Attending to mathematical language with transformers
Artit Wangperawong. 2018 · 2018
Later among the works it cites.
Wikiqa: A challenge dataset for open-domain question answering
Yi Yang, Wen-tau Yih, and Christopher Meek. 2015 · 2018
Later among the works it cites.
Quaternion convolutional neural networks
Xuanyu Zhu, Yi Xu, Hongteng Xu, and Changjian Chen. 2018 · 2018
Later among the works it cites.
Phrase-based attentions
Phi Xuan Nguyen and Shafiq Joty. 2019 · 2019
Closest in time.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Closest in time.
Complex embeddings for simple link prediction
Théo Trouillon, Johannes Welbl, Sebastian Riedel, Éric Gaussier, and Guillaume Bouchard. 2016 · 2080
Closest in time.