Fetching the paper…
Reading the bibliography…
The pre-trained language model (eg, BERT) based deep retrieval models achieved superior performance over lexical retrieval models (eg, BM25) in many passage retrieval tasks.
Okapi at trec-3
Stephen E. Robertson, Steve Walker, Susan Jones, Micheline Hancock-Beaulieu, and Mike Gatford · 1994
Earlier work this paper cites.
A language modeling approach to information retrieval
Jay M. Ponte and W. Bruce Croft · 1998
Earlier work this paper cites.
Bridging the lexical chasm: Statistical approaches to answer-finding
Adam Berger, Rich Caruana, David Cohn, Dayne Freitag, and Vibhu Mittal · 2000
Earlier work this paper cites.
Relevance based language models
Victor Lavrenko and W. Bruce Croft · 2001
Earlier work this paper cites.
Probabilistic models of information retrieval based on measuring the divergence from randomness
Gianni Amati and Cornelis Joost Van Rijsbergen · 2002
Earlier work this paper cites.
Umass at TREC 2004: Novelty and HARD
Nasreen Abdul Jaleel, James Allan, W. Bruce Croft, Fernando Diaz, Leah S. Larkey, Xiaoyan Li, Mark D. Smucker, and Courtney Wade · 2004
Earlier work this paper cites.
Overview of the trec 2004 robust retrieval track, 2005-08-01 2005
Ellen Voorhees · 2005
Earlier work this paper cites.
Language model information retrieval with document expansion
Tao Tao, Xuanhui Wang, Qiaozhu Mei, and ChengXiang Zhai · 2006
Earlier work this paper cites.
Reciprocal rank fusion outperforms condorcet and individual rank learning methods
Gordon V. Cormack, Charles L A Clarke, and Stefan Buettcher · 2009
Earlier work this paper cites.
From puppy to maturity: Experiences in developing terrier
Craig Macdonald, Richard McCreadie, Rodrygo LT Santos, and Iadh Ounis · 2012
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, undefinedukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset, 2018
Payal Bajaj, Daniel Campos, Nick Craswell, Li Deng, Jianfeng Gao, Xiaodong Liu, Rangan Majumder, Andrew McNamara, Bhaskar Mitra, Tri Nguyen, Mir Rosenberg, Xia Song, Alina Stoica, Saurabh Tiwary, and Tong Wang · 2018
Earlier work this paper cites.
Hierarchical neural story generation
Angela Fan, Mike Lewis, and Yann N. Dauphin · 2018
Earlier work this paper cites.
Anserini: Reproducible ranking baselines using lucene
Peilin Yang, Hui Fang, and Jimmy Lin · 2018
Earlier work this paper cites.
Context-aware sentence/passage term importance estimation for first stage retrieval
Zhuyun Dai and Jamie Callan · 2019
Earlier work this paper cites.
Deeper text understanding for ir with contextual neural language modeling
Zhuyun Dai and Jamie Callan · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
From doc2query to doctttttquery
Rodrigo Nogueira and Jimmy Lin · 2019
Cited alongside, same era.
Document expansion by query prediction
Rodrigo Nogueira, Wei Yang, Jimmy Lin, and Kyunghyun Cho · 2019
Cited alongside, same era.
RRF102: meeting the TREC-COVID challenge with a 100+ runs ensemble
Michael Bendersky, Honglei Zhuang, Ji Ma, Shuguang Han, Keith B. Hall, and Ryan T. McDonald · 2020
Cited alongside, same era.
Unsupervised corpus aware language model pre-training for dense passage retrieval
Luyu Gao and Jamie Callan · 2021
Later among the works it cites.
Complement lexical retrieval model with semantic residual embeddings
Luyu Gao, Zhuyun Dai, Tongfei Chen, Zhen Fan, Benjamin Van Durme, and Jamie Callan · 2021
Later among the works it cites.
Billion-scale similarity search with gpus
Jeff Johnson, Matthijs Douze, and Hervé Jégou · 2021
Later among the works it cites.
Jimmy Lin and Xueguang Ma · 2021
Later among the works it cites.
In-batch negatives for knowledge distillation with tightly-coupled teachers for dense retrieval
Sheng-Chieh Lin, Jheng-Hong Yang, and Jimmy Lin · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Orcas: 20 million clicked query-document pairs for analyzing search
Nick Craswell, Daniel Campos, Bhaskar Mitra, Emine Yilmaz, and Bodo Billerbeck · 2020
Cited alongside, same era.
Context-aware term weighting for first stage passage retrieval
Zhuyun Dai and Jamie Callan · 2020
Cited alongside, same era.
Accelerating large-scale inference with anisotropic vector quantization
Ruiqi Guo, Philip Sun, Erik Lindgren, Quan Geng, David Simcha, Felix Chern, and Sanjiv Kumar · 2020
Cited alongside, same era.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih · 2020
Cited alongside, same era.
Saar Kuzi, Mingyang Zhang, Cheng Li, Michael Bendersky, and Marc Najork · 2020
Cited alongside, same era.
BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2020
Cited alongside, same era.
Multi-stage training with improved negative contrast for neural passage retrieval
Jing Lu, Gustavo Hernández Ábrego, Ji Ma, Jianmo Ni, and Yinfei Yang · 2021
Later among the works it cites.
Sparse, Dense, and Attentional Representations for Text Retrieval
Yi Luan, Jacob Eisenstein, Kristina Toutanova, and Michael Collins · 2021
Later among the works it cites.
A replication study of dense passage retriever
Xueguang Ma, Kai Sun, Ronak Pradeep, and Jimmy Lin · 2021
Later among the works it cites.
Generation-augmented retrieval for open-domain question answering
Yuning Mao, Pengcheng He, Xiaodong Liu, Yelong Shen, Jianfeng Gao, Jiawei Han, and Weizhu Chen · 2021
Later among the works it cites.
The expando-mono-duo design pattern for text ranking with pretrained sequence-to-sequence models
Ronak Pradeep, Rodrigo Nogueira, and Jimmy Lin · 2021
Later among the works it cites.
RocketQA: An optimized training approach to dense passage retrieval for open-domain question answering
Yingqi Qu, Yuchen Ding, Jing Liu, Kai Liu, Ruiyang Ren, Wayne Xin Zhao, Daxiang Dong, Hua Wu, and Haifeng Wang · 2021
Later among the works it cites.
Simple entity-centric questions challenge dense retrievers
Christopher Sciavolino, Zexuan Zhong, Jinhyuk Lee, and Danqi Chen · 2021
Later among the works it cites.
BEIR: A heterogeneous benchmark for zero-shot evaluation of information retrieval models
Nandan Thakur, Nils Reimers, Andreas Rücklé, Abhishek Srivastava, and Iryna Gurevych · 2021
Later among the works it cites.
Bert-based dense retrievers require interpolation with bm25 for effective passage retrieval
Shuai Wang, Shengyao Zhuang, and Guido Zuccon · 2021
Later among the works it cites.
Approximate nearest neighbor negative contrastive learning for dense text retrieval
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul N. Bennett, Junaid Ahmed, and Arnold Overwijk · 2021
Later among the works it cites.