Fetching the paper…
Reading the bibliography…
Neural rankers based on deep pretrained language models (LMs) have been shown to improve many information retrieval benchmarks.
The vocabulary problem in human-system communication
George W. Furnas, Thomas K. Landauer, Louis M. Gomez, and Susan T. Dumais. 1987 · 1987
Earlier work this paper cites.
Improving retrieval performance by relevance feedback
Gerard Salton and Chris Buckley. 1990 · 1990
Earlier work this paper cites.
Search engines: Information retrieval in practice . Vol. 520
W Bruce Croft, Donald Metzler, and Trevor Strohman. 2010 · 2010
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2015 · 2015
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset
Payal Bajaj, Daniel Campos, Nick Craswell, Li Deng, Jianfeng Gao, Xiaodong Liu, Rangan Majumder, Andrew McNamara, Bhaskar Mitra, Tri Nguyen, et al · 2016
Earlier work this paper cites.
Deeprank: A new deep architecture for relevance ranking in information retrieval. In Proceedings of CIKM . 257–266
Liang Pang, Yanyan Lan, Jiafeng Guo, Jun Xu, Jingfang Xu, and Xueqi Cheng. 2017 · 2017
Earlier work this paper cites.
Anserini: Enabling the use of Lucene for information retrieval research. In Proceedings of SIGIR . 1253–1256
Peilin Yang, Hui Fang, and Jimmy Lin. 2017 · 2017
Earlier work this paper cites.
Convolutional neural networks for soft-matching n-grams in ad-hoc search. In Proceedings of WSDM . 126–134
Zhuyun Dai, Chenyan Xiong, Jamie Callan, and Zhiyuan Liu. 2018 · 2018
Earlier work this paper cites.
SciBERT: A Pretrained Language Model for Scientific Text. In Proceedings of EMNLP-IJCNLP . 3606–3611
Iz Beltagy, Kyle Lo, and Arman Cohan. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of NAACL . 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
On the effect of low-frequency terms on neural-IR models. In Proceedings of SIGIR . 1137–1140
Sebastian Hofstätter, Navid Rekabsaz, Carsten Eickhoff, and Allan Hanbury. 2019 · 2019
Cited alongside, same era.
CEDR: Contextualized embeddings for document ranking. In Proceedings of SIGIR . 1101–1104
Sean MacAvaney, Andrew Yates, Arman Cohan, and Nazli Goharian. 2019 · 2019
Cited alongside, same era.
Rodrigo Nogueira and Kyunghyun Cho. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Cited alongside, same era.
Simple applications of BERT for ad hoc document retrieval
Wei Yang, Haotian Zhang, and Jimmy Lin. 2019 · 2019
Cited alongside, same era.
Dense Passage Retrieval for Open-Domain Question Answering
Vladimir Karpukhin, Barlas Oğuz, Sewon Min, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020 · 2020
Closest in time.
BioBERT: a pre-trained biomedical language representation model for biomedical text mining
Jinhyuk Lee, Wonjin Yoon, Sungdong Kim, Donghyeon Kim, Sunkyu Kim, Chan Ho So, and Jaewoo Kang. 2020 · 2020
Closest in time.
Sparse, Dense, and Attentional Representations for Text Retrieval
Yi Luan, Jacob Eisenstein, Kristina Toutanova, and Michael Collins. 2020 · 2020
Closest in time.
Zero-shot Neural Retrieval via Domain-targeted Synthetic Query Generation
Ji Ma, Ivan Korotkov, Yinfei Yang, Keith Hall, and Ryan McDonald. 2020 · 2020
Closest in time.
SLEDGE: A Simple Yet Effective Baseline for Coronavirus Scientific Knowledge Search
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Generic intent representation in web search. In Proceedings of SIGIR . 65–74
Hongfei Zhang, Xia Song, Chenyan Xiong, Corby Rosset, Paul N Bennett, Nick Craswell, and Saurabh Tiwary. 2019 · 2019
Cited alongside, same era.
Pre-training tasks for embedding-based large-scale retrieval
Wei-Cheng Chang, Felix X Yu, Yin-Wen Chang, Yiming Yang, and Sanjiv Kumar. 2020 · 2020
Cited alongside, same era.
Overview of the trec 2019 deep learning track
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Campos, and Ellen M Voorhees. 2020 · 2020
Cited alongside, same era.
Complementing Lexical Retrieval with Semantic Residual Embedding
Luyu Gao, Zhuyun Dai, Zhen Fan, and Jamie Callan. 2020 · 2020
Cited alongside, same era.
Don’t Stop Pretraining: Adapt Language Models to Domains and Tasks
Suchin Gururangan, Ana Marasović, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, and Noah A Smith. 2020 · 2020
Cited alongside, same era.
Sean MacAvaney, Arman Cohan, and Nazli Goharian. 2020 · 2020
Closest in time.
TREC-COVID: Constructing a Pandemic Information Retrieval Test Collection
Ellen Voorhees, Tasmeer Alam, Steven Bedrick, Dina Demner-Fushman, William R Hersh, Kyle Lo, Kirk Roberts, Ian Soboroff, and Lucy Lu Wang. 2020 · 2020
Closest in time.
CORD-19: The Covid-19 Open Research Dataset
Lucy Lu Wang, Kyle Lo, Yoganand Chandrasekhar, Russell Reas, Jiangjiang Yang, Darrin Eide, Kathryn Funk, Rodney Kinney, Ziyang Liu, William Merrill, et al · 2020
Closest in time.
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul Bennett, Junaid Ahmed, and Arnold Overwijk. 2020 · 2020
Closest in time.
Selective Weak Supervision for Neural Information Retrieval. In Proceedings of WWW . 474–485
Kaitao Zhang, Chenyan Xiong, Zhenghao Liu, and Zhiyuan Liu. 2020 · 2020
Closest in time.