Fetching the paper…
Reading the bibliography…
Knowledge-intensive language tasks (KILT) usually require a large body of information to provide correct answers.
A vector space model for automatic indexing
Gerard Salton, Anita Wong, and Chung-Shu Yang. 1975 · 1975
Earlier work this paper cites.
Relevance weighting of search terms
Stephen E Robertson and K Sparck Jones. 1976 · 1976
Earlier work this paper cites.
Learning term discrimination. In SIGIR . 1993–1996
Jibril Frej, Philippe Mulhem, Didier Schwab, and Jean-Pierre Chevallet. 2020 · 1996
Earlier work this paper cites.
Learning to rank with nonsmooth cost functions
Christopher Burges, Robert Ragno, and Quoc Le. 2006 · 2006
Earlier work this paper cites.
Learning to rank for information retrieval
Tie-Yan Liu et al · 2009
Earlier work this paper cites.
The probabilistic relevance framework: BM25 and beyond
Stephen Robertson and Hugo Zaragoza. 2009 · 2009
Earlier work this paper cites.
Robust disambiguation of named entities in text. In EMNLP . 782–792
Johannes Hoffart, Mohamed Amir Yosef, Ilaria Bordino, Hagen Fürstenau, Manfred Pinkal, Marc Spaniol, Bilyana Taneva, Stefan Thater, and Gerhard Weikum. 2011 · 2011
Earlier work this paper cites.
Generating text with recurrent neural networks. In ICML
Ilya Sutskever, James Martens, and Geoffrey E Hinton. 2011 · 2011
Earlier work this paper cites.
Learning to rank for information retrieval and natural language processing
Hang Li. 2014 · 2014
Earlier work this paper cites.
Learning to reweight terms with distributed representations. In SIGIR . 575–584
Guoqing Zheng and Jamie Callan. 2015 · 2015
Earlier work this paper cites.
A deep relevance matching model for ad-hoc retrieval. In CIKM . 55–64
Jiafeng Guo, Yixing Fan, Qingyao Ai, and W Bruce Croft. 2016 · 2016
Earlier work this paper cites.
Reading Wikipedia to answer open-domain questions. In 55th Annual Meeting of the Association for Computational Linguistics, ACL 2017 . 1870–1879
Danqi Chen, Adam Fisch, Jason Weston, and Antoine Bordes. 2017 · 2017
Earlier work this paper cites.
TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension. In ACL . Association for Computational Linguistics, Vancouver, Canada, 1601–1611
Mandar Joshi, Eunsol Choi, Daniel Weld, and Luke Zettlemoyer. 2017 · 2017
Earlier work this paper cites.
Zero-Shot Relation Extraction via Reading Comprehension. In CoNLL . 333–342
Omer Levy, Minjoon Seo, Eunsol Choi, and Luke Zettlemoyer. 2017 · 2017
Earlier work this paper cites.
Convolutional neural networks for soft-matching n-grams in ad-hoc search. In WSDM . 126–134
Zhuyun Dai, Chenyan Xiong, Jamie Callan, and Zhiyuan Liu. 2018 · 2018
Earlier work this paper cites.
Wizard of Wikipedia: Knowledge-Powered Conversational Agents. In International Conference on Learning Representations
Emily Dinan, Stephen Roller, Kurt Shuster, Angela Fan, Michael Auli, and Jason Weston. 2018 · 2018
Earlier work this paper cites.
T-rex: A large scale alignment of natural language with knowledge base triples. In LREC
Hady Elsahar, Pavlos Vougiouklis, Arslen Remaci, Christophe Gravier, Jonathon Hare, Frederique Laforest, and Elena Simperl. 2018 · 2018
Cited alongside, same era.
Robust named entity disambiguation with random walks
Zhaochen Guo and Denilson Barbosa. 2018 · 2018
Cited alongside, same era.
FEVER: a Large-scale Dataset for Fact Extraction and VERification. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) . 809–819
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Cited alongside, same era.
HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering. In EMNLP . Association for Computational Linguistics, Brussels, Belgium, 2369–2380
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018 · 2018
Cited alongside, same era.
Scalable Zero-shot Entity Linking with Dense Entity Retrieval. In EMNLP . 6397–6407
Ledell Wu, Fabio Petroni, Martin Josifoski, Sebastian Riedel, and Luke Zettlemoyer. 2020 · 2020
Later among the works it cites.
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval. In ICLR
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul N Bennett, Junaid Ahmed, and Arnold Overwijk. 2020 · 2020
Later among the works it cites.
Pegasus: Pre-training with extracted gap-sentences for abstractive summarization. In International Conference on Machine Learning . PMLR, 11328–11339
Jingqing Zhang, Yao Zhao, Mohammad Saleh, and Peter Liu. 2020 · 2020
Later among the works it cites.
SimCSE: Simple Contrastive Learning of Sentence Embeddings. In EMNLP . 6894–6910
Tianyu Gao, Xingcheng Yao, and Danqi Chen. 2021 · 2021
Later among the works it cites.
Efficiently teaching an effective dense retriever with balanced topic aware sampling. In SIGIR . 113–122
Sebastian Hofstätter, Sheng-Chieh Lin, Jheng-Hong Yang, Jimmy Lin, and Allan Hanbury. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
FLAIR: An easy-to-use framework for state-of-the-art NLP. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics (Demonstrations) . 54–59
Alan Akbik, Tanja Bergmann, Duncan Blythe, Kashif Rasul, Stefan Schweter, and Roland Vollgraf. 2019 · 2019
Cited alongside, same era.
Pre-training Tasks for Embedding-based Large-scale Retrieval. In ICLR
Wei-Cheng Chang, X Yu Felix, Yin-Wen Chang, Yiming Yang, and Sanjiv Kumar. 2019 · 2019
Cited alongside, same era.
ELI5: Long Form Question Answering. In ACL . Association for Computational Linguistics, Florence, Italy, 3558–3567
Angela Fan, Yacine Jernite, Ethan Perez, David Grangier, Jason Weston, and Michael Auli. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In NAACL-HLT . 4171–4186
Jacob Devlin Ming-Wei Chang Kenton and Lee Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Natural questions: a benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, et al · 2019
Cited alongside, same era.
Latent Retrieval for Weakly Supervised Open Domain Question Answering. In ACL . 6086–6096
Kenton Lee, Ming-Wei Chang, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Context-aware term weighting for first stage passage retrieval. In SIGIR . 1533–1536
Zhuyun Dai and Jamie Callan. 2020 · 2020
Cited alongside, same era.
Autoregressive Entity Retrieval. In ICLR
Nicola De Cao, Gautier Izacard, Sebastian Riedel, and Fabio Petroni. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
Multi-Task Retrieval for Knowledge-Intensive Tasks. In ACL . 1098–1111
Jean Maillard, Vladimir Karpukhin, Fabio Petroni, Wen-tau Yih, Barlas Oguz, Veselin Stoyanov, and Gargi Ghosh. 2021 · 2021
Later among the works it cites.
Rethinking search: making domain experts out of dilettantes. In ACM SIGIR Forum , Vol. 55. ACM New York, NY, USA, 1–27
Donald Metzler, Yi Tay, Dara Bahri, and Marc Najork. 2021 · 2021
Later among the works it cites.
KILT: a Benchmark for Knowledge Intensive Language Tasks. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, Online, 2523–2544
Fabio Petroni, Aleksandra Piktus, Angela Fan, Patrick Lewis, Majid Yazdani, Nicola De Cao, James Thorne, Yacine Jernite, Vladimir Karpukhin, Jean Maillard, Vassilis Plachouras, Tim Rocktäschel, and Sebastian Riedel. 2021 · 2021
Later among the works it cites.
Optimizing dense retrieval model training with hard negatives. In SIGIR . 1503–1512
Jingtao Zhan, Jiaxin Mao, Yiqun Liu, Jiafeng Guo, Min Zhang, and Shaoping Ma. 2021 · 2021
Later among the works it cites.
Michele Bevilacqua, Giuseppe Ottaviano, Patrick Lewis, Wen tau Yih, Sebastian Riedel, and Fabio Petroni. 2022 · 2022
Closest in time.
GERE: Generative Evidence Retrieval for Fact Verification
Jiangui Chen, Ruqing Zhang, Jiafeng Guo, Yixing Fan, and Xueqi Cheng. 2022 · 2022
Closest in time.
Semantic models for the first-stage retrieval: A comprehensive review
Jiafeng Guo, Yinqiong Cai, Yixing Fan, Fei Sun, Ruqing Zhang, and Xueqi Cheng. 2022 · 2022
Closest in time.
TABi: Type-Aware Bi-Encoders for Open-Domain Entity Retrieval
Megan Leszczynski, Daniel Y Fu, Mayee F Chen, and Christopher Ré. 2022 · 2022
Closest in time.
Pre-train a Discriminative Text Encoder for Dense Retrieval via Contrastive Span Prediction
Xinyu Ma, Jiafeng Guo, Ruqing Zhang, Yixing Fan, and Xueqi Cheng. 2022 · 2022
Closest in time.
Transformer memory as a differentiable search index
Yi Tay, Vinh Q Tran, Mostafa Dehghani, Jianmo Ni, Dara Bahri, Harsh Mehta, Zhen Qin, Kai Hui, Zhe Zhao, Jai Gupta, et al · 2022
Closest in time.
DynamicRetriever: A Pre-training Model-based IR System with Neither Sparse nor Dense Index
Yujia Zhou, Jing Yao, Zhicheng Dou, Ledell Wu, and Ji-Rong Wen. 2022 · 2022
Closest in time.