Fetching the paper…
Reading the bibliography…
Despite Portuguese being one of the most spoken languages in the world, there is a lack of high-quality information retrieval datasets in that language.
Cross-language information retrieval (clir) track overview
Peter Schäuble and Páraic Sheridan · 1998
Earlier work this paper cites.
The importance of evaluation for cross-language system development: the clef experience
Carol Peters and Martin Braschler · 2002
Earlier work this paper cites.
Rcv1: A new benchmark collection for text categorization research
David D Lewis, Yiming Yang, Tony Russell-Rose, and Fan Li · 2004
Earlier work this paper cites.
Relevance assessment: are judges exchangeable and does it matter
Peter Bailey, Nick Craswell, Ian Soboroff, Paul Thomas, Arjen P de Vries, and Emine Yilmaz · 2008
Earlier work this paper cites.
Reciprocal rank fusion outperforms condorcet and individual rank learning methods
Gordon V Cormack, Charles LA Clarke, and Stefan Buettcher · 2009
Earlier work this paper cites.
Fasttext.zip: Compressing text classification models
Armand Joulin, Edouard Grave, Piotr Bojanowski, Matthijs Douze, Hérve Jégou, and Tomas Mikolov · 2016
Earlier work this paper cites.
Bag of tricks for efficient text classification
Armand Joulin, Edouard Grave, Piotr Bojanowski, and Tomas Mikolov · 2016
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset
Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng · 2016
Earlier work this paper cites.
Gauging the quality of relevance assessments using inter-rater agreement
Tadele T Damessie, Thao P Nghiem, Falk Scholer, and J Shane Culpepper · 2017
Earlier work this paper cites.
doccano: Text annotation tool for human, 2018
Hiroki Nakayama, Takahiro Kubo, Junya Kamura, Yasufumi Taniguchi, and Xu Liang · 2018
Earlier work this paper cites.
Billion-scale similarity search with GPUs
Jeff Johnson, Matthijs Douze, and Hervé Jégou · 2019
Earlier work this paper cites.
Tydi qa: A benchmark for information-seeking question answering in ty pologically di verse languages
Jonathan H Clark, Eunsol Choi, Michael Collins, Dan Garrette, Tom Kwiatkowski, Vitaly Nikolaev, and Jennimaria Palomaki · 2020
Earlier work this paper cites.
mt5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel · 2020
Cited alongside, same era.
mmarco: A multilingual version of the ms marco passage ranking dataset
Luiz Bonifacio, Vitor Jeronymo, Hugo Queiroz Abonizio, Israel Campiotti, Marzieh Fadaee, Roberto Lotufo, and Rodrigo Nogueira · 2021
Cited alongside, same era.
Trec deep learning track: Reusable test collections in the large data regime
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Campos, Ellen M Voorhees, and Ian Soboroff · 2021
Cited alongside, same era.
Splade v2: Sparse lexical and expansion model for information retrieval
Thibault Formal, Carlos Lassance, Benjamin Piwowarski, and Stéphane Clinchant · 2021
Cited alongside, same era.
Regis: A test collection for geoscientific documents in portuguese
Clueweb22: 10 billion web documents with rich information
Arnold Overwijk, Chenyan Xiong, and Jamie Callan · 2022
Later among the works it cites.
Too many relevants: Whither cranfield test collections?
Ellen M Voorhees, Nick Craswell, and Jimmy Lin · 2022
Later among the works it cites.
Can old trec collections reliably evaluate modern neural retrieval models?
Ellen M Voorhees, Ian Soboroff, and Jimmy Lin · 2022
Later among the works it cites.
Text embeddings by weakly-supervised contrastive pre-training
Liang Wang, Nan Yang, Xiaolong Huang, Binxing Jiao, Linjun Yang, Daxin Jiang, Rangan Majumder, and Furu Wei · 2022
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, brian ichter, Fei Xia, Ed Chi, Quoc V Le, and Denny Zhou · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lucas Lima de Oliveira, Regis Kruel Romeu, and Viviane Pereira Moreira · 2021
Cited alongside, same era.
Pyserini: A Python toolkit for reproducible information retrieval research with sparse and dense representations
Jimmy Lin, Xueguang Ma, Sheng-Chieh Lin, Jheng-Hong Yang, Ronak Pradeep, and Rodrigo Nogueira · 2021
Cited alongside, same era.
Scaling language models: Methods, analysis & insights from training gopher
Jack W Rae, Sebastian Borgeaud, Trevor Cai, Katie Millican, Jordan Hoffmann, Francis Song, John Aslanides, Sarah Henderson, Roman Ring, Susannah Young, et al · 2021
Cited alongside, same era.
Evaluating Information Retrieval and Access Tasks: NTCIR’s Legacy of Research Impact
Tetsuya Sakai, Douglas W Oard, and Noriko Kando · 2021
Cited alongside, same era.
mrobust04: A multilingual version of the trec robust 2004 benchmark
Vitor Jeronymo, Mauricio Nascimento, Roberto Lotufo, and Rodrigo Nogueira · 2022
Cited alongside, same era.
Hc4: A new suite of test collections for ad hoc clir
Dawn Lawrie, James Mayfield, Douglas W Oard, and Eugene Yang · 2022
Cited alongside, same era.
Transfer learning approaches for building cross-language dense retrieval models
Suraj Nair, Eugene Yang, Dawn Lawrie, Kevin Duh, Paul McNamee, Kenton Murray, James Mayfield, and Douglas W Oard · 2022
Cited alongside, same era.
Perspectives on large language models for relevance judgment
Guglielmo Faggioli, Laura Dietz, Charles Clarke, Gianluca Demartini, Matthias Hagen, Claudia Hauff, Noriko Kando, Evangelos Kanoulas, Martin Potthast, Benno Stein, et al · 2023
Later among the works it cites.
Large language models can accurately predict searcher preferences
Paul Thomas, Seth Spielman, Nick Craswell, and Bhaskar Mitra · 2023
Later among the works it cites.
M3exam: A multilingual, multimodal, multilevel benchmark for examining large language models
Wenxuan Zhang, Sharifah Mahani Aljunied, Chang Gao, Yew Ken Chia, and Lidong Bing · 2023
Later among the works it cites.
Miracl: A multilingual retrieval dataset covering 18 diverse languages
Xinyu Zhang, Nandan Thakur, Odunayo Ogundepo, Ehsan Kamalloo, David Alfonso-Hermelo, Xiaoguang Li, Qun Liu, Mehdi Rezagholizadeh, and Jimmy Lin · 2023
Later among the works it cites.
An exam-based evaluation approach beyond traditional relevance judgments
Naghmeh Farzi and Laura Dietz · 2024
Closest in time.
Enhancing human annotation: Leveraging large language models and efficient batch processing
Oleg Zendel, J Shane Culpepper, Falk Scholer, and Paul Thomas · 2024
Closest in time.