Fetching the paper…
Reading the bibliography…
Robust 2004 is an information retrieval benchmark whose large number of judgments per query make it a reliable evaluation dataset.
Cross-language information retrieval (clir) track overview
P. Schäuble and P. Sheridan · 1998
Earlier work this paper cites.
Clef 2001 — overview of results
M. Braschler · 2001
Earlier work this paper cites.
European research letter: Cross-language system evaluation: The clef campaigns
C. Peters and M. Braschler · 2001
Earlier work this paper cites.
Clef 2002—overview of results
M. Braschler · 2002
Earlier work this paper cites.
Clef 2003–overview of results
M. Braschler · 2003
Earlier work this paper cites.
Overview of the TREC 2004 robust track
E. M. Voorhees · 2004
Earlier work this paper cites.
Alternatives to bpref
T. Sakai · 2007
Earlier work this paper cites.
Overview of fire 2008
M. Mitra · 2008
Cited alongside, same era.
Natural language processing with Python: analyzing text with the natural language toolkit
S. Bird, E. Klein, and E. Loper · 2009
Cited alongside, same era.
Overview of fire 2010
P. Majumder, D. Pal, A. Bandyopadhyay, and M. Mitra · 2013
Cited alongside, same era.
spaCy 2: Natural language understanding with Bloom embeddings, convolutional neural networks and incremental parsing
M. Honnibal and I. Montani · 2017
Cited alongside, same era.
MS MARCO: A Human Generated MAchine Reading COmprehension Dataset
P. Bajaj, D. Campos, N. Craswell, L. Deng, J. Gao, X. Liu, R. Majumder, A. McNamara, B. Mitra, T. Nguyen, M. Rosenberg, X. Song, A. Stoica, S. Tiwary, and T. Wang · 2018
Cited alongside, same era.
mmarco: A multilingual version of MS MARCO passage ranking dataset
Pretrained transformers for text ranking: Bert and beyond
J. Lin, R. Nogueira, and A. Yates · 2021
Later among the works it cites.
Evaluating Information Retrieval and Access Tasks: NTCIR’s Legacy of Research Impact
T. Sakai, D. W. Oard, and N. Kando · 2021
Later among the works it cites.
Mr. tydi: A multi-lingual benchmark for dense retrieval
X. Zhang, X. Ma, P. Shi, and J. Lin · 2021
Later among the works it cites.
Hc4: a new suite of test collections for ad hoc clir
D. Lawrie, J. Mayfield, D. W. Oard, and E. Yang · 2022
Closest in time.
No parameter left behind: How distillation and model size affect zero-shot retrieval
G. M. Rosa, L. Bonifacio, V. Jeronymo, H. Abonizio, M. Fadaee, R. Lotufo, and R. Nogueira · 2022
Closest in time.
Too many relevants: Whither cranfield test collections?
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
L. H. Bonifacio, V. Jeronymo, H. Q. Abonizio, I. Campiotti, R. de Alencar Lotufo, and R. Nogueira · 2021
Cited alongside, same era.
Trec deep learning track: reusable test collections in the large data regime
N. Craswell, B. Mitra, E. Yilmaz, D. Campos, E. M. Voorhees, and I. Soboroff · 2021
Cited alongside, same era.
Can old trec collections reliably evaluate modern neural retrieval models?
E. M. Voorhees, I. Soboroff, and J. Lin
Cited in the paper.
E. M. Voorhees, N. Craswell, and J. Lin · 2022
Closest in time.