Fetching the paper…
Reading the bibliography…
We present a memory-augmented approach to condition an autoregressive language model on a knowledge graph.
Leveraging knowledge bases in lstms for improving machine reading
Bishan Yang and Tom M. Mitchell. 2019 · 1902
Earlier work this paper cites.
Dynamic evaluation of transformer language models
Ben Krause, Emmanuel Kahembwe, Iain Murray, and Steve Renals. 2019 · 1904
Earlier work this paper cites.
ERNIE: enhanced representation through knowledge integration
Yu Sun, Shuohuan Wang, Yu-Kun Li, Shikun Feng, Xuyi Chen, Han Zhang, Xin Tian, Danxiang Zhu, Hao Tian, and Hua Wu. 2019 · 1904
Earlier work this paper cites.
KG-BERT: BERT for knowledge graph completion
Liang Yao, Chengsheng Mao, and Yuan Luo. 2019 · 1909
Earlier work this paper cites.
Non-parametric adaptation for neural machine translation
Ankur Bapna and Orhan Firat. 2019 · 1931
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey E. Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Freebase: A shared database of structured general human knowledge
Kurt D. Bollacker, Robert P. Cook, and Patrick Tufts. 2007 · 1963
Earlier work this paper cites.
Interpolated estimation of markov source parameters from sparse data
Frederick Jelinek. 1980 · 1980
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
REALM: retrieval-augmented language model pre-training
Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Ming-Wei Chang. 2020 · 2002
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Janvin. 2003 · 2003
Earlier work this paper cites.
Using tf-idf to determine word relevance in document queries
Juan Ramos et al. 2003 · 2003
Earlier work this paper cites.
Modeling local coherence: An entity-based approach
Regina Barzilay and Mirella Lapata. 2005 · 2005
Earlier work this paper cites.
Language models are few-shot learners
Tom B Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 2005
Earlier work this paper cites.
A survey of named entity recognition and classification
David Nadeau and Satoshi Sekine. 2007 · 2007
Earlier work this paper cites.
Open information extraction from the web
Oren Etzioni, Michele Banko, Stephen Soderland, and Daniel S Weld. 2008 · 2008
Earlier work this paper cites.
Word meaning in minds and machines
Brenden M. Lake and Gregory L. Murphy. 2020 · 2008
Earlier work this paper cites.
Design challenges and misconceptions in named entity recognition
Lev Ratinov and Dan Roth. 2009 · 2009
Earlier work this paper cites.
Nearest neighbor machine translation
Urvashi Khandelwal, Angela Fan, Dan Jurafsky, Luke Zettlemoyer, and Mike Lewis. 2020a · 2010
Earlier work this paper cites.
Language models are open knowledge graphs
Chenguang Wang, Xiao Liu, and Dawn Song. 2020 · 2010
Earlier work this paper cites.
Multi-task feature learning for knowledge graph enhanced recommendation
Hongwei Wang, Fuzheng Zhang, Miao Zhao, Wenjie Li, Xing Xie, and Minyi Guo. 2019 · 2010
Earlier work this paper cites.
Thinking, Fast and Slow
Daniel Kahneman. 2011 · 2011
Earlier work this paper cites.
The human knowledge compression contest
Marcus Hutter. 2012 · 2012
Earlier work this paper cites.
Translating embeddings for modeling multi-relational data
Antoine Bordes, Nicolas Usunier, Alberto García-Durán, Jason Weston, and Oksana Yakhnenko. 2013 · 2013
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches
KyungHyun Cho, Bart van Merrienboer, Dzmitry Bahdanau, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Leveraging linguistic structure for open domain information extraction
Gabor Angeli, Melvin Jose Johnson Premkumar, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Learning knowledge graphs for question answering through conversational dialog
Ben Hixon, Peter Clark, and Hannaneh Hajishirzi. 2015 · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
A neural knowledge language model
Sungjin Ahn, Heeyoul Choi, Tanel Pärnamaa, and Yoshua Bengio. 2016 · 2016
Earlier work this paper cites.
Improving neural language models with a continuous cache
Edouard Grave, Armand Joulin, and Nicolas Usunier. 2016 · 2016
Cited alongside, same era.
Gaussian error linear units (gelus)
Dan Hendrycks and Kevin Gimpel. 2016 · 2016
Cited alongside, same era.
Tying word vectors and word classifiers: A loss framework for language modeling
Hakan Inan, Khashayar Khosravi, and Richard Socher. 2016 · 2016
Cited alongside, same era.
Globally coherent text generation with neural checklist models
Chloé Kiddon, Luke Zettlemoyer, and Yejin Choi. 2016 · 2016
Cited alongside, same era.
YAGO: A multilingual knowledge base from wikipedia, wordnet, and geonames
Thomas Rebele, Fabian M. Suchanek, Johannes Hoffart, Joanna Biega, Erdal Kuzey, and Gerhard Weikum. 2016 · 2016
Knowledge graph embedding based question answering
Xiao Huang, Jingyuan Zhang, Dingcheng Li, and Ping Li. 2019 · 2019
Later among the works it cites.
Linguistic knowledge and transferability of contextual representations
Nelson F. Liu, Matt Gardner, Yonatan Belinkov, Matthew E. Peters, and Noah A. Smith. 2019 · 2019
Later among the works it cites.
Barack’s wife hillary: Using knowledge graphs for fact-aware language modeling
Robert L. Logan, Nelson F. Liu, Matthew E. Peters, Matt Gardner, and Sameer Singh. 2019 · 2019
Later among the works it cites.
Episodic memory in lifelong language learning
Cyprien de Masson d’Autume, Sebastian Ruder, Lingpeng Kong, and Dani Yogatama. 2019 · 2019
Later among the works it cites.
Opendialkg: Explainable conversational reasoning with attention-based walks over knowledge graphs
Seungwhan Moon, Pararth Shah, Anuj Kumar, and Rajen Subba. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Collaborative knowledge base embedding for recommender systems
Fuzheng Zhang, Nicholas Jing Yuan, Defu Lian, Xing Xie, and Wei-Ying Ma. 2016 · 2016
Cited alongside, same era.
Learning to compute word embeddings on the fly
Dzmitry Bahdanau, Tom Bosc, Stanislaw Jastrzebski, Edward Grefenstette, Pascal Vincent, and Yoshua Bengio. 2017 · 2017
Cited alongside, same era.
Reading wikipedia to answer open-domain questions
Danqi Chen, Adam Fisch, Jason Weston, and Antoine Bordes. 2017 · 2017
Cited alongside, same era.
Efficient softmax approximation for GPUs
Édouard Grave, Armand Joulin, Moustapha Cissé, David Grangier, and Hervé Jégou. 2017 · 2017
Cited alongside, same era.
Dynamic entity representations in neural language models
Yangfeng Ji, Chenhao Tan, Sebastian Martschat, Yejin Choi, and Noah A. Smith. 2017 · 2017
Cited alongside, same era.
Pointer sentinel mixture models
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Malte Ostendorff, Peter Bourgonje, Maria Berger, Julián Moreno Schneider, Georg Rehm, and Bela Gipp. 2019 · 2019
Later among the works it cites.
Knowledge enhanced contextual word representations
Matthew E. Peters, Mark Neumann, Robert L. Logan IV, Roy Schwartz, Vidur Joshi, Sameer Singh, and Noah A. Smith. 2019 · 2019
Later among the works it cites.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick S. H. Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander H. Miller. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Quaternion knowledge graph embeddings
Shuai Zhang, Yi Tay, Lina Yao, and Qi Liu. 2019 · 2019
Later among the works it cites.
Latent relation language models
Hiroaki Hayashi, Zecong Hu, Chenyan Xiong, and Graham Neubig. 2020 · 2020
Later among the works it cites.
Haiku: Sonnet for JAX
Tom Hennigan, Trevor Cai, Tamara Norman, and Igor Babuschkin. 2020 · 2020
Later among the works it cites.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick S. H. Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020 · 2020
Later among the works it cites.
Generalization through memorization: Nearest neighbor language models
Urvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer, and Mike Lewis. 2020b · 2020
Later among the works it cites.
Retrieval-augmented generation for knowledge-intensive NLP tasks
Patrick S. H. Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela. 2020 · 2020
Later among the works it cites.
K-BERT: enabling language representation with knowledge graph
Weijie Liu, Peng Zhou, Zhe Zhao, Zhiruo Wang, Qi Ju, Haotang Deng, and Ping Wang. 2020 · 2020
Later among the works it cites.
Differentiable reasoning on large knowledge bases and natural language
Pasquale Minervini, Matko Bosnjak, Tim Rocktäschel, Sebastian Riedel, and Edward Grefenstette. 2020 · 2020
Later among the works it cites.
Knowledge graph based synthetic corpus generation for knowledge-enhanced language model pre-training
Oshin Agarwal, Heming Ge, Siamak Shakeri, and Rami Al-Rfou. 2021 · 2021
Later among the works it cites.
Autoregressive entity retrieval
Nicola De Cao, Gautier Izacard, Sebastian Riedel, and Fabio Petroni. 2021 · 2021
Later among the works it cites.
Augmenting transformers with knn-based composite memory for dialog
Angela Fan, Claire Gardent, Chloé Braud, and Antoine Bordes. 2021 · 2021
Later among the works it cites.
Leveraging passage retrieval with generative models for open domain question answering
Gautier Izacard and Edouard Grave. 2021 · 2021
Later among the works it cites.
Truthfulqa: Measuring how models mimic human falsehoods
Stephanie Lin, Jacob Hilton, and Owain Evans. 2021 · 2021
Later among the works it cites.
Maxwell Nye, Michael Henry Tessler, Joshua B. Tenenbaum, and Brenden M. Lake. 2021 · 2021
Later among the works it cites.
Efficient retrieval augmented generation from unstructured knowledge for task-oriented dialog
David Thulke, Nico Daheim, Christian Dugast, and Hermann Ney. 2021 · 2021
Later among the works it cites.
Adaptable and interpretable neural memoryover symbolic knowledge
Pat Verga, Haitian Sun, Livio Baldini Soares, and William W. Cohen. 2021 · 2021
Later among the works it cites.
QA-GNN: reasoning with language models and knowledge graphs for question answering
Michihiro Yasunaga, Hongyu Ren, Antoine Bosselut, Percy Liang, and Jure Leskovec. 2021 · 2021
Later among the works it cites.
Adaptive semiparametric language models
Dani Yogatama, Cyprien de Masson d’Autume, and Lingpeng Kong. 2021 · 2021
Later among the works it cites.
Complex embeddings for simple link prediction
Théo Trouillon, Johannes Welbl, Sebastian Riedel, Éric Gaussier, and Guillaume Bouchard. 2016 · 2080
Closest in time.