Fetching the paper…
Reading the bibliography…
This paper describes our participation in the 2023 WSDM CUP - MIRACL challenge.
Colbert: Efficient and effective passage search via contextualized late interaction over bert
O. Khattab and M. Zaharia · 2020
Earlier work this paper cites.
Document ranking with a pretrained sequence-to-sequence model
R. Nogueira, Z. Jiang, and J. Lin · 2020
Earlier work this paper cites.
mmarco: A multilingual version of the ms marco passage ranking dataset
L. Bonifacio, V. Jeronymo, H. Q. Abonizio, I. Campiotti, M. Fadaee, R. Lotufo, and R. Nogueira · 2021
Earlier work this paper cites.
Rethink training of bert rerankers in multi-stage retrieval pipeline
L. Gao, Z. Dai, and J. Callan · 2021
Earlier work this paper cites.
Debertav3: Improving deberta using electra-style pre-training with gradient-disentangled embedding sharing, 2021
P. He, J. Gao, and W. Chen · 2021
Earlier work this paper cites.
Towards unsupervised dense information retrieval with contrastive learning
G. Izacard, M. Caron, L. Hosseini, S. Riedel, P. Bojanowski, A. Joulin, and E. Grave · 2021
Earlier work this paper cites.
ranx.fuse: A python library for metasearch
E. Bassani and L. Romelli · 2022
Cited alongside, same era.
Scaling instruction-finetuned language models, 2022
H. W. Chung, L. Hou, S. Longpre, B. Zoph, Y. Tay, W. Fedus, E. Li, X. Wang, M. Dehghani, S. Brahma, A. Webson, S. S. Gu, Z. Dai, M. Suzgun, X. Chen, A. Chowdhery, S. Narang, G. Mishra, A. Yu, V. Zhao, Y. Huang, A. Dai, H. Yu, S. Petrov, E. H. Chi, J. Dean, J. Devlin, A. Roberts, D. Zhou, Q. V. Le, and J. Wei · 2022
Cited alongside, same era.
No language left behind: Scaling human-centered machine translation
M. R. Costa-jussà, J. Cross, O. Çelebi, M. Elbayad, K. Heafield, K. Heffernan, E. Kalbassi, J. Lam, D. Licht, J. Maillard, et al · 2022
Cited alongside, same era.
From distillation to hard negative sampling: Making sparse neural ir models more effective
T. Formal, C. Lassance, B. Piwowarski, and S. Clinchant · 2022
Cited alongside, same era.
An efficiency study for splade models
C. Lassance and S. Clinchant · 2022
Cited alongside, same era.
Crosslingual generalization through multitask finetuning
N. Muennighoff, T. Wang, L. Sutawika, A. Roberts, S. Biderman, T. L. Scao, M. S. Bari, S. Shen, Z.-X. Yong, H. Schoelkopf, et al · 2022
Later among the works it cites.
Byt5: Towards a token-free future with pre-trained byte-to-byte models
L. Xue, A. Barua, N. Constant, R. Al-Rfou, S. Narang, M. Kale, A. Roberts, and C. Raffel · 2022
Later among the works it cites.
Making a miracl: Multilingual information retrieval across a continuum of languages
X. Zhang, N. Thakur, O. Ogundepo, E. Kamalloo, D. Alfonso-Hermelo, X. Li, Q. Liu, M. Rezagholizadeh, and J. Lin · 2022
Later among the works it cites.
Rankt5: Fine-tuning t5 for text ranking with ranking losses
H. Zhuang, Z. Qin, R. Jagerman, K. Hui, J. Ma, J. Lu, J. Ni, X. Wang, and M. Bendersky · 2022
Later among the works it cites.
An experimental study on pretraining transformers from scratch for ir
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Lassance, H. Déjean, and S. Clinchant · 2023
Closest in time.