Fetching the paper…
Reading the bibliography…
A retrieval model should not only interpolate the training data but also extrapolate well to the queries that are different from the training data.
A Coefficient of Agreement for Nominal Scales
Jacob Cohen. 1960 · 1960
Earlier work this paper cites.
Aslib–Cranfield research project
Cyril Cleverdon and EM Keen. 1966 · 1966
Earlier work this paper cites.
Relevance weighting of search terms
Stephen E. Robertson and Karen Spärck Jones. 1976 · 1976
Earlier work this paper cites.
Extrapolation and interpolation in neural network classifiers
Etienne Barnard and LFA Wessels. 1992 · 1992
Earlier work this paper cites.
Extrapolation limitations of multilayer feedforward neural networks. In [Proceedings 1992] IJCNN International Joint Conference on Neural Networks , Vol. 4. IEEE, 25–30
Pamela J Haley and DONALD Soloway. 1992 · 1992
Earlier work this paper cites.
Cumulated gain-based evaluation of IR techniques
Kalervo Järvelin and Jaana Kekäläinen. 2002 · 2002
Earlier work this paper cites.
Rank-biased precision for measurement of retrieval effectiveness
Alistair Moffat and Justin Zobel. 2008 · 2008
Earlier work this paper cites.
User variability and IR system evaluation. In Proceedings of The 38th International ACM SIGIR conference on research and development in Information Retrieval . 625–634
Peter Bailey, Alistair Moffat, Falk Scholer, and Paul Thomas. 2015 · 2015
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset
Payal Bajaj, Daniel Campos, Nick Craswell, Li Deng, Jianfeng Gao, Xiaodong Liu, Rangan Majumder, Andrew McNamara, Bhaskar Mitra, Tri Nguyen, et al · 2016
Earlier work this paper cites.
Building machines that learn and think like people
Brenden M Lake, Tomer D Ullman, Joshua B Tenenbaum, and Samuel J Gershman. 2017 · 2017
Earlier work this paper cites.
Evaluating web search with a bejeweled player model. In Proceedings of the 40th International ACM SIGIR Conference on Research and Development in Information Retrieval . 425–434
Fan Zhang, Yiqun Liu, Xin Li, Min Zhang, Yinghui Xu, and Shaoping Ma. 2017 · 2017
Earlier work this paper cites.
Measuring the utility of search engine result pages: an information foraging based measure. In The 41st international acm sigir conference on research & development in information retrieval . 605–614
Leif Azzopardi, Paul Thomas, and Nick Craswell. 2018 · 2018
Earlier work this paper cites.
WWW’18 Open Challenge: Financial Opinion Mining and Question Answering
Macedo Maia, Siegfried Handschuh, André Freitas, Brian Davis, Ross McDermott, Manel Zarrouk, and Alexandra Balahur. 2018 · 2018
Earlier work this paper cites.
Fever: a large-scale dataset for fact extraction and verification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Earlier work this paper cites.
Retrieval of the best counterargument without prior topic knowledge. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 241–251
Henning Wachsmuth, Shahbaz Syed, and Benno Stein. 2018 · 2018
Earlier work this paper cites.
Anserini: Reproducible ranking baselines using Lucene
Peilin Yang, Hui Fang, and Jimmy Lin. 2018 · 2018
Earlier work this paper cites.
Overview of the TREC 2019 deep learning track. In Text REtrieval Conference (TREC) . TREC
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Campos, and Ellen M. Voorhees. 2020a · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
RoBERTa: a robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 2019
Earlier work this paper cites.
Overview of the NTCIR-14 we want web task
Jiaxin Mao, Tetsuya Sakai, Cheng Luo, Peng Xiao, Yiqun Liu, and Zhicheng Dou. 2019 · 2019
Cited alongside, same era.
Rodrigo Nogueira and Kyunghyun Cho. 2019 · 2019
Cited alongside, same era.
Shiori Sagawa, Pang Wei Koh, Tatsunori B Hashimoto, and Percy Liang. 2019 · 2019
Cited alongside, same era.
Modeling generalization in machine learning: A methodological and computational study
Pietro Barbiero, Giovanni Squillero, and Alberto Tonda. 2020 · 2020
Cited alongside, same era.
Overview of Touché 2020: Argument Retrieval. In CLEF
Alexander Bondarenko, Maik Fröbe, Meriem Beloucif, Lukas Gienapp, Yamen Ajjour, Alexander Panchenko, Christian Biemann, Benno Stein, Henning Wachsmuth, Martin Potthast, and Matthias Hagen. 2020 · 2020
Unsupervised corpus aware language model pre-training for dense passage retrieval
Luyu Gao and Jamie Callan. 2021b · 2021
Later among the works it cites.
Efficiently Teaching an Effective Dense Retriever with Balanced Topic Aware Sampling. In Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR ’21) . 113–122
Sebastian Hofstätter, Sheng-Chieh Lin, Jheng-Hong Yang, Jimmy Lin, and Allan Hanbury. 2021 · 2021
Later among the works it cites.
Out-of-distribution generalization via risk extrapolation (rex). In International Conference on Machine Learning . PMLR, 5815–5826
David Krueger, Ethan Caballero, Joern-Henrik Jacobsen, Amy Zhang, Jonathan Binas, Dinghuai Zhang, Remi Le Priol, and Aaron Courville. 2021 · 2021
Later among the works it cites.
Jimmy Lin and Xueguang Ma. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Overview of the TREC 2020 Deep Learning Track
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Fernando Campos, and Ellen M. Voorhees. 2020b · 2020
Cited alongside, same era.
Climate-fever: A dataset for verification of real-world climate claims
Thomas Diggelmann, Jordan Boyd-Graber, Jannis Bulian, Massimiliano Ciaramita, and Markus Leippold. 2020 · 2020
Cited alongside, same era.
Realm: retrieval-augmented language model pre-training
Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Ming-Wei Chang. 2020 · 2020
Cited alongside, same era.
Dense Passage Retrieval for Open-Domain Question Answering. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP)
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020 · 2020
Cited alongside, same era.
ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERT
O. Khattab and M. Zaharia. 2020 · 2020
Cited alongside, same era.
Distilling Dense Representations for Ranking using Tightly-Coupled Teachers
Sheng-Chieh Lin, Jheng-Hong Yang, and Jimmy Lin. 2020 · 2020
Cited alongside, same era.
Overview of the NTCIR-15 we want web with CENTRE (WWW-3) task
Tetsuya Sakai, Sijie Tao, Zhaohao Zeng, Yukun Zheng, Jiaxin Mao, Zhumin Chu, Yiqun Liu, Maria Maistro, Zhicheng Dou, Nicola Ferro, et al · 2020
Cited alongside, same era.
Pretrained Transformers for Text Ranking: BERT and Beyond
Jimmy J. Lin, Rodrigo Nogueira, and Andrew Yates. 2021 · 2021
Later among the works it cites.
Pre-trained language model for web-scale retrieval in baidu search. In Proceedings of the 27th ACM SIGKDD Conference on Knowledge Discovery & Data Mining . 3365–3375
Yiding Liu, Weixue Lu, Suqi Cheng, Daiting Shi, Shuaiqiang Wang, Zhicong Cheng, and Dawei Yin. 2021 · 2021
Later among the works it cites.
Less is More: Pretrain a Strong Siamese Encoder for Dense Text Retrieval Using a Weak Decoder. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . 2780–2791
Shuqi Lu, Chenyan Xiong, Di He, Guolin Ke, Waleed Malik, Zhicheng Dou, Paul Bennett, Tieyan Liu, and Arnold Overwijk. 2021 · 2021
Later among the works it cites.
Learning Passage Impacts for Inverted Indexes. In Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR ’21)
Antonio Mallia, Omar Khattab, Torsten Suel, and Nicola Tonellotto. 2021 · 2021
Later among the works it cites.
RocketQA: An optimized training approach to dense passage retrieval for open-domain question answering. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . 5835–5847
Yingqi Qu, Yuchen Ding, Jing Liu, Kai Liu, Ruiyang Ren, Wayne Xin Zhao, Daxiang Dong, Hua Wu, and Haifeng Wang. 2021 · 2021
Later among the works it cites.
ColBERTv2: Effective and Efficient Retrieval via Lightweight Late Interaction
Keshav Santhanam, Omar Khattab, Jon Saad-Falcon, Christopher Potts, and Matei Zaharia. 2021 · 2021
Later among the works it cites.
Simple Entity-Centric Questions Challenge Dense Retrievers. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . 6138–6148
Christopher Sciavolino, Zexuan Zhong, Jinhyuk Lee, and Danqi Chen. 2021 · 2021
Later among the works it cites.
BEIR: A Heterogeneous Benchmark for Zero-shot Evaluation of Information Retrieval Models. In Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 2)
Nandan Thakur, Nils Reimers, Andreas Rücklé, Abhishek Srivastava, and Iryna Gurevych. 2021 · 2021
Later among the works it cites.
Zero-Shot Dense Retrieval with Momentum Adversarial Domain Invariant Representations
Ji Xin, Chenyan Xiong, Ashwin Srinivasan, Ankita Sharma, Damien Jose, and Paul N Bennett. 2021 · 2021
Later among the works it cites.
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval. In International Conference on Learning Representations
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul N. Bennett, Junaid Ahmed, and Arnold Overwijk. 2021 · 2021
Later among the works it cites.
Optimizing Dense Retrieval Model Training with Hard Negatives. In Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR ’21) . 1503–1512
Jingtao Zhan, Jiaxin Mao, Yiqun Liu, Jiafeng Guo, Min Zhang, and Shaoping Ma. 2021 · 2021
Later among the works it cites.
An online learning approach to interpolation and extrapolation in domain generalization. In International Conference on Artificial Intelligence and Statistics . PMLR, 2641–2657
Elan Rosenfeld, Pradeep Ravikumar, and Andrej Risteski. 2022 · 2022
Closest in time.
Exploring the robust extrapolation of high-dimensional machine learning potentials
Claudio Zeni, Andrea Anelli, Aldo Glielmo, and Kevin Rossi. 2022 · 2022
Closest in time.
Learning Discrete Representations via Constrained Clustering for Effective and Efficient Dense Retrieval. In Proceedings of the Fifteenth ACM International Conference on Web Search and Data Mining . 1328–1336
Jingtao Zhan, Jiaxin Mao, Yiqun Liu, Jiafeng Guo, Min Zhang, and Shaoping Ma. 2022 · 2022
Closest in time.