Fetching the paper…
Reading the bibliography…
We introduce HUBERT which combines the structured-representational power of Tensor-Product Representations (TPRs) and BERT, a pre-trained bidirectional Transformer language model.
Unifying question answering and text classification via span extraction
Nitish Shirish Keskar, Bryan McCann, Caiming Xiong, and Richard Socher. 2019 · 1904
Earlier work this paper cites.
BERT rediscovers the classical NLP pipeline
Ian Tenney, Dipanjan Das, and Ellie Pavlick. 2019 · 1905
Earlier work this paper cites.
Visualizing and measuring the geometry of BERT
Andy Coenen, Emily Reif, Ann Yuan, Been Kim, Adam Pearce, Fernanda Viégas, and Martin Wattenberg. 2019 · 1906
Earlier work this paper cites.
Open sesame: Getting inside BERT’s linguistic knowledge
Yongjie Lin, Yi Chern Tan, and Robert Frank. 2019 · 1906
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019b · 1907
Earlier work this paper cites.
Semantics-aware bert for language understanding
Zhuosheng Zhang, Yuwei Wu, Hai Zhao, Zuchao Li, Shuailiang Zhang, Xi Zhou, and Xiang Zhou. 2019 · 1909
Earlier work this paper cites.
Physical symbol systems
Allen Newell. 1980 · 1980
Earlier work this paper cites.
Tensor product variable binding and the representation of symbolic structures in connectionist systems
Paul Smolensky. 1990 · 1990
Earlier work this paper cites.
Holographic reduced representations
Tony A Plate. 1995 · 1995
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
The Harmonic Mind: From Neural Computation to Optimality-Theoretic GrammarVolume I: Cognitive Architecture (Bradford Books)
Paul Smolensky and Géraldine Legendre. 2006 · 2006
Earlier work this paper cites.
Vector symbolic architectures: A new building material for artificial general intelligence
Simon D Levy and Ross Gayler. 2008 · 2008
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
D. P. Kingma and J. Ba. 2014 · 2014
Earlier work this paper cites.
Sequence to Sequence Learning with Neural Networks
I. Sutskever, O. Vinyals, and Q. V. Le. 2014 · 2014
Cited alongside, same era.
A large annotated corpus for learning natural language inference
Samuel R Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning. 2015 · 2015
Cited alongside, same era.
Deep Residual Learning for Image Recognition
K. He, X. Zhang, S. Ren, and J. Sun. 2015 · 2015
Cited alongside, same era.
A Simple Way to Initialize Recurrent Networks of Rectified Linear Units
Q. V. Le, N. Jaitly, and G. E. Hinton. 2015 · 2015
Cited alongside, same era.
Reasoning in vector space: An exploratory study of question answering
Moontae Lee, Xiaodong He, Wen-tau Yih, Jianfeng Gao, Li Deng, and Paul Smolensky. 2016 · 2016
Tensor product generation networks for deep NLP modeling
Qiuyuan Huang, Paul Smolensky, Xiaodong He, Li Deng, and Dapeng Oliver Wu. 2018 · 2018
Later among the works it cites.
Question-answering with grammatically-interpretable representations
Hamid Palangi, Paul Smolensky, Xiaodong He, and Li Deng. 2018 · 2018
Later among the works it cites.
Sentence encoders on stilts: Supplementary training on intermediate labeled-data tasks
Jason Phang, Thibault Févry, and Samuel R Bowman. 2018 · 2018
Later among the works it cites.
Learning to reason with third order tensor products
Imanol Schlag and Jürgen Schmidhuber. 2018 · 2018
Later among the works it cites.
Glue: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R Bowman. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
J. Lei Ba, J. R. Kiros, and G. E. Hinton. 2016 · 2016
Cited alongside, same era.
SQuAD: 100,000+ Questions for Machine Comprehension of Text
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang. 2016 · 2016
Cited alongside, same era.
Squad: 100, 000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Cited alongside, same era.
Towards ai-complete question answering: A set of prerequisite toy tasks
Jason Weston, Antoine Bordes, Sumit Chopra, and Tomas Mikolov. 2016 · 2016
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
On orthogonality and learning recurrent networks with long term dependencies
E. Vorontsov, C. Trabelsi, S. Kadoury, and C. Pal. 2017 · 2017
Cited alongside, same era.
Can we gain more from orthogonality regularizations in training deep cnns?
Nitin Bansal, Xiaohan Chen, and Zhangyang Wang. 2018 · 2018
Cited alongside, same era.
What Does BERT Look At? An Analysis of BERT’s Attention
K. Clark, U. Khandelwal, O. Levy, and C. D. Manning. 2019 · 2019
Closest in time.
A structural probe for finding syntax in word representations
John Hewitt and Christopher D Manning. 2019 · 2019
Closest in time.
Revealing the Dark Secrets of BERT
O. Kovaleva, A. Romanov, A. Rogers, and A. Rumshisky. 2019 · 2019
Closest in time.
Right for the wrong reasons: Diagnosing syntactic heuristics in natural language inference
Tom McCoy, Ellie Pavlick, and Tal Linzen. 2019 · 2019
Closest in time.
Are Sixteen Heads Really Better than One?
P. Michel, O. Levy, and G. Neubig. 2019 · 2019
Closest in time.
Can you tell me how to get past sesame street? sentence-level pretraining beyond language modeling
Alex Wang, Jan Hula, Patrick Xia, Raghavendra Pappagari, R Thomas McCoy, Roma Patel, Najoung Kim, Ian Tenney, Yinghui Huang, Katherin Yu, et al. 2019 · 2019
Closest in time.
BERTs of a feather do not generalize together: Large variability in generalization across models with similar test set performance
R. Thomas McCoy, Junghyun Min, and Tal Linzen. 2020 · 2020
Closest in time.