Fetching the paper…
Reading the bibliography…
Pre-trained language models are increasingly important components across multiple information retrieval (IR) paradigms.
Rodrigo Nogueira and Kyunghyun Cho. 2019 · 1901
Earlier work this paper cites.
Okapi at TREC-3
Stephen E Robertson, Steve Walker, Susan Jones, Micheline M Hancock-Beaulieu, Mike Gatford, et al · 1995
Earlier work this paper cites.
Query evaluation: strategies and optimizations
Howard Turtle and James Flood. 1995 · 1995
Earlier work this paper cites.
Efficient query evaluation using a two-level retrieval process. In CIKM
Andrei Z Broder, David Carmel, Michael Herscovici, Aya Soffer, and Jason Zien. 2003 · 2003
Earlier work this paper cites.
Improving Efficient Neural Ranking Models with Cross-Architecture Knowledge Distillation
Sebastian Hofstätter, Sophia Althammer, Michael Schröder, Mete Sertkan, and Allan Hanbury. 2020 · 2010
Earlier work this paper cites.
Product quantization for nearest neighbor search
Herve Jegou, Matthijs Douze, and Cordelia Schmid. 2010 · 2010
Earlier work this paper cites.
Learning To Retrieve: How to Train a Dense Retrieval Model Effectively and Efficiently
Jingtao Zhan, Jiaxin Mao, Yiqun Liu, Min Zhang, and Shaoping Ma. 2020 · 2010
Earlier work this paper cites.
Faster top-k document retrieval using block-max indexes. In SIGIR
Shuai Ding and Torsten Suel. 2011 · 2011
Earlier work this paper cites.
Optimizing top-k document retrieval strategies for block-max indexes. In WSDM
Constantinos Dimopoulos, Sergey Nepomnyachiy, and Torsten Suel. 2013 · 2013
Earlier work this paper cites.
MS MARCO: A Human-Generated MAchine Reading COmprehension Dataset
Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng. 2016 · 2016
Earlier work this paper cites.
Faster blockmax wand with variable-sized blocks. In SIGIR
Antonio Mallia, Giuseppe Ottaviano, Elia Porciani, Nicola Tonellotto, and Rossano Venturini. 2017 · 2017
Earlier work this paper cites.
Efficient and robust approximate nearest neighbor search using hierarchical navigable small world graphs
Yu A Malkov and Dmitry A Yashunin. 2018 · 2018
Earlier work this paper cites.
Efficient Query Processing for Scalable Web Search
Nicola Tonellotto, Craig Macdonald, Iadh Ounis, et al · 2018
Earlier work this paper cites.
Anserini: Reproducible ranking baselines using Lucene
Peilin Yang, Hui Fang, and Jimmy Lin. 2018 · 2018
Earlier work this paper cites.
To index or not to index: Optimizing exact maximum inner product search. In 2019 IEEE 35th International Conference on Data Engineering (ICDE) . IEEE, 1250–1261
Firas Abuzaid, Geet Sethi, Peter Bailis, and Matei Zaharia. 2019 · 2019
Earlier work this paper cites.
Billion-scale similarity search with gpus
Jeff Johnson, Matthijs Douze, and Hervé Jégou. 2019 · 2019
Earlier work this paper cites.
Natural Questions: A Benchmark for Question Answering Research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, Kristina Toutanova, Llion Jones, Matthew Kelcey, Ming-Wei Chang, Andrew M. Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov. 2019 · 2019
Earlier work this paper cites.
CEDR: Contextualized Embeddings for Document Ranking. In Proceedings of the 42nd International ACM SIGIR Conference on Research and Development in Information Retrieval . 1101–1104
Sean MacAvaney, Andrew Yates, Arman Cohan, and Nazli Goharian. 2019 · 2019
Earlier work this paper cites.
PISA: Performant indexes and search for academia
Antonio Mallia, Michal Siedlaczek, Joel Mackenzie, and Torsten Suel. 2019 · 2019
Cited alongside, same era.
Context-Aware Term Weighting For First Stage Passage Retrieval. In Proceedings of the 43rd International ACM SIGIR conference on research and development in Information Retrieval, SIGIR 2020, Virtual Event, China, July 25-30, 2020 , Jimmy Huang, Yi Chang, Xueqi Cheng, Jaap Kamps, Vanessa Murdock, Ji-Rong Wen, and Yiqun Liu (Eds.). ACM, 1533–1536
Zhuyun Dai and Jamie Callan. 2020 · 2020
Cited alongside, same era.
Poly-encoders: Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence Scoring. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26-30, 2020 . OpenReview.net
Samuel Humeau, Kurt Shuster, Marie-Anne Lachaux, and Jason Weston. 2020 · 2020
Cited alongside, same era.
Dense Passage Retrieval for Open-Domain Question Answering. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) . Association for Computational Linguistics, Online, 6769–6781
Learning passage impacts for inverted indexes. In Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval . 1723–1727
Antonio Mallia, Omar Khattab, Torsten Suel, and Nicola Tonellotto. 2021 · 2021
Later among the works it cites.
RocketQA: An Optimized Training Approach to Dense Passage Retrieval for Open-Domain Question Answering. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, Online, 5835–5847
Yingqi Qu, Yuchen Ding, Jing Liu, Kai Liu, Ruiyang Ren, Wayne Xin Zhao, Daxiang Dong, Hua Wu, and Haifeng Wang. 2021 · 2021
Later among the works it cites.
RocketQAv2: A Joint Training Method for Dense Passage Retrieval and Passage Re-ranking
Ruiyang Ren, Yingqi Qu, Jing Liu, Wayne Xin Zhao, Qiaoqiao She, Hua Wu, Haifeng Wang, and Ji-Rong Wen. 2021 · 2021
Later among the works it cites.
ColBERTv2: Effective and Efficient Retrieval via Lightweight Late Interaction
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020 · 2020
Cited alongside, same era.
Finding the best of both worlds: Faster and more robust top-k document retrieval. In Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval . 1031–1040
Omar Khattab, Mohammad Hammoud, and Tamer Elsayed. 2020 · 2020
Cited alongside, same era.
ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERT. In Proceedings of the 43rd International ACM SIGIR conference on research and development in Information Retrieval, SIGIR 2020, Virtual Event, China, July 25-30, 2020 , Jimmy Huang, Yi Chang, Xueqi Cheng, Jaap Kamps, Vanessa Murdock, Ji-Rong Wen, and Yiqun Liu (Eds.). ACM, 39–48
Omar Khattab and Matei Zaharia. 2020 · 2020
Cited alongside, same era.
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval. In International Conference on Learning Representations
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul N Bennett, Junaid Ahmed, and Arnold Overwijk. 2020 · 2020
Cited alongside, same era.
Passage Collection (Augmented)
Anserini GitHub Repo Authors. 2021 · 2021
Cited alongside, same era.
Pretrained Transformer Language Models for Search - part 3
Jo Kristian Bergum. 2021 · 2021
Cited alongside, same era.
Overview of the TREC 2021 deep learning track. In Text REtrieval Conference (TREC) . TREC
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Campos, and Jimmy Lin. 2022 · 2021
Cited alongside, same era.
SPLADE v2: Sparse Lexical and Expansion Model for Information Retrieval
Thibault Formal, Carlos Lassance, Benjamin Piwowarski, and Stéphane Clinchant. 2021a · 2021
Cited alongside, same era.
COIL: Revisit Exact Lexical Match in Information Retrieval with Contextualized Inverted List. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, Online, 3030–3042
Luyu Gao, Zhuyun Dai, and Jamie Callan. 2021 · 2021
Cited alongside, same era.
Keshav Santhanam, Omar Khattab, Jon Saad-Falcon, Christopher Potts, and Matei Zaharia. 2021 · 2021
Later among the works it cites.
BEIR: A Heterogeneous Benchmark for Zero-shot Evaluation of Information Retrieval Models. In Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 2)
Nandan Thakur, Nils Reimers, Andreas Rücklé, Abhishek Srivastava, and Iryna Gurevych. 2021 · 2021
Later among the works it cites.
Query embedding pruning for dense retrieval. In Proceedings of the 30th ACM International Conference on Information & Knowledge Management . 3453–3457
Nicola Tonellotto and Craig Macdonald. 2021 · 2021
Later among the works it cites.
Pseudo-relevance feedback for multiple representation dense retrieval. In Proceedings of the 2021 ACM SIGIR International Conference on Theory of Information Retrieval . 297–306
Xiao Wang, Craig Macdonald, Nicola Tonellotto, and Iadh Ounis. 2021 · 2021
Later among the works it cites.
Anserini Regressions: MS MARCO (V2) Passage Ranking
Anserini GitHub Repo Authors. 2022 · 2022
Closest in time.
Sebastian Hofstätter, Omar Khattab, Sophia Althammer, Mete Sertkan, and Allan Hanbury. 2022 · 2022
Closest in time.
A proposed conceptual framework for a representational approach to information retrieval. In ACM SIGIR Forum , Vol. 55. ACM New York, NY, USA, 1–29
Jimmy Lin. 2022 · 2022
Closest in time.
Toward A Fine-Grained Analysis of Distribution Shifts in MSMARCO
Simon Lupart and Stéphane Clinchant. 2022 · 2022
Closest in time.
Zero-shot Query Contextualization for Conversational Search
Antonios Minas Krasakis, Andrew Yates, and Evangelos Kanoulas. 2022 · 2022
Closest in time.
Hindsight: Posterior-guided Training of Retrievers for Improved Open-ended Generation. In International Conference on Learning Representations
Ashwin Paranjape, Omar Khattab, Christopher Potts, Matei Zaharia, and Christopher D Manning. 2022 · 2022
Closest in time.
RELIC: Retrieving Evidence for Literary Claims
Katherine Thai, Yapei Chang, Kalpesh Krishna, and Mohit Iyyer. 2022 · 2022
Closest in time.
Curriculum Learning for Dense Retrieval Distillation
Hansi Zeng, Hamed Zamani, and Vishwa Vinay. 2022 · 2022
Closest in time.
Evaluating Extrapolation Performance of Dense Retrieval
Jingtao Zhan, Xiaohui Xie, Jiaxin Mao, Yiqun Liu, Min Zhang, and Shaoping Ma. 2022 · 2022
Closest in time.
Evaluating Token-Level and Passage-Level Dense Retrieval Models for Math Information Retrieval
Wei Zhong, Jheng-Hong Yang, and Jimmy Lin. 2022 · 2022
Closest in time.