Fetching the paper…
Reading the bibliography…
Large-scale test collections play a crucial role in Information Retrieval (IR) research.
The aslib cranfield research project on the comparative efficiency of indexing systems. In Aslib Proceedings , Vol. 12. 421–431
Cyril W Cleverdon. 1960 · 1960
Earlier work this paper cites.
Passage Retrieval Revisited. In Proceedings of SIGIR
Marcin Kaszkiel and Justin Zobel. 1997 · 1997
Earlier work this paper cites.
The TREC robust retrieval track. In ACM SIGIR Forum , Vol. 39. ACM New York, NY, USA, 11–20
Ellen M Voorhees. 2005 · 2005
Earlier work this paper cites.
Passage retrieval and evaluation
Courtney Wade and James Allan. 2005 · 2005
Earlier work this paper cites.
Evaluation over thousands of queries. In Proceedings of the 31st annual international ACM SIGIR conference on Research and development in information retrieval . 651–658
Ben Carterette, Virgil Pavlu, Evangelos Kanoulas, Javed A Aslam, and James Allan. 2008 · 2008
Earlier work this paper cites.
Overview of the TREC 2009 Web Track.. In Trec , Vol. 9. 20–29
Charles LA Clarke, Nick Craswell, and Ian Soboroff. 2009 · 2009
Earlier work this paper cites.
Test collection based evaluation of information retrieval systems
Mark Sanderson et al · 2010
Earlier work this paper cites.
Improving test collection pools with machine learning. In Proceedings of the 19th Australasian Document Computing Symposium . 2–9
Gaya K Jayasinghe, William Webber, Mark Sanderson, and J Shane Culpepper. 2014 · 2014
Earlier work this paper cites.
MS MARCO: A Human Generated MAchine Reading COmprehension Dataset. In Proceedings of the Workshop on Cognitive Computation: Integrating neural and symbolic approaches 2016 co-located with the 30th Annual Conference on Neural Information Processing Systems (NIPS 2016), Barcelona, Spain, December 9, 2016
Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng. 2016 · 2016
Cited alongside, same era.
Overview of the TREC 2019 deep learning track. In Trec
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Campos, and Ellen M Voorhees. 2020 · 2019
Cited alongside, same era.
Simplified data wrangling with ir_datasets. In Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval . 2429–2436
Sean MacAvaney, Andrew Yates, Sergey Feldman, Doug Downey, Arman Cohan, and Nazli Goharian. 2021 · 2021
Cited alongside, same era.
BEIR: A Heterogeneous Benchmark for Zero-shot Evaluation of Information Retrieval Models. In Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 2)
Nandan Thakur, Nils Reimers, Andreas Rücklé, Abhishek Srivastava, and Iryna Gurevych. 2021 · 2021
G-Eval: NLG Evaluation using Gpt-4 with Better Human Alignment. In The 2023 Conference on Empirical Methods in Natural Language Processing
Yang Liu, Dan Iter, Yichong Xu, Shuohang Wang, Ruochen Xu, and Chenguang Zhu. 2023 · 2023
Later among the works it cites.
Large language models can accurately predict searcher preferences
Paul Thomas, Seth Spielman, Nick Craswell, and Bhaskar Mitra. 2023 · 2023
Later among the works it cites.
Query2doc: Query Expansion with Large Language Models. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 9414–9423
Liang Wang, Nan Yang, and Furu Wei. 2023 · 2023
Later among the works it cites.
Synthetic Test Collections for Retrieval Evaluation
Hossein A Rahmani, Nick Craswell, Emine Yilmaz, Bhaskar Mitra, and Daniel Campos. 2024a · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Cited alongside, same era.
Task-aware Retrieval with Instructions. In Findings of the Association for Computational Linguistics: ACL 2023 . 3650–3675
Akari Asai, Timo Schick, Patrick Lewis, Xilun Chen, Gautier Izacard, Sebastian Riedel, Hannaneh Hajishirzi, and Wen-tau Yih. 2023 · 2023
Cited alongside, same era.
Overview of the TREC 2023 Deep Learning Track. In Text REtrieval Conference (TREC) . NIST, TREC
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Hossein A. Rahmani, Daniel Campos, Jimmy Lin, Ellen M. Voorhees, and Ian Soboroff. 2024 · 2023
Cited alongside, same era.
Perspectives on large language models for relevance judgment. In Proceedings of the 2023 ACM SIGIR International Conference on Theory of Information Retrieval . 39–50
Guglielmo Faggioli, Laura Dietz, Charles LA Clarke, Gianluca Demartini, Matthias Hagen, Claudia Hauff, Noriko Kando, Evangelos Kanoulas, Martin Potthast, Benno Stein, et al · 2023
Cited alongside, same era.
LLM4Eval: Large Language Model for Evaluation in IR. In Proceedings of the 47th International ACM SIGIR Conference on Research and Development in Information Retrieval (Washington DC, USA) (SIGIR ’24) . Association for Computing Machinery, New York, NY, USA, 3040–3043
Hossein A. Rahmani, Clemencia Siro, Mohammad Aliannejadi, Nick Craswell, Charles L. A. Clarke, Guglielmo Faggioli, Bhaskar Mitra, Paul Thomas, and Emine Yilmaz. 2024c
Cited in the paper.
Hossein A Rahmani, Clemencia Siro, Mohammad Aliannejadi, Nick Craswell, Charles LA Clarke, Guglielmo Faggioli, Bhaskar Mitra, Paul Thomas, and Emine Yilmaz. 2024b · 2024
Closest in time.
LLMJudge: LLMs for Relevance Judgments
Hossein A Rahmani, Emine Yilmaz, Nick Craswell, Bhaskar Mitra, Paul Thomas, Charles LA Clarke, Mohammad Aliannejadi, Clemencia Siro, and Guglielmo Faggioli. 2024d · 2024
Closest in time.
Dense text retrieval based on pretrained language models: A survey
Wayne Xin Zhao, Jing Liu, Ruiyang Ren, and Ji-Rong Wen. 2024 · 2024
Closest in time.