Fetching the paper…
Reading the bibliography…
The LLMJudge challenge is organized as part of the LLM4Eval workshop at SIGIR 2024.
Perspectives on large language models for relevance judgment,
G. Faggioli, L. Dietz, C. L. Clarke, G. Demartini, M. Hagen, C. Hauff, N. Kando, E. Kanoulas, M. Potthast, B. Stein, et al., · 2023
Earlier work this paper cites.
Large language models can accurately predict searcher preferences,
P. Thomas, S. Spielman, N. Craswell, B. Mitra, · 2023
Earlier work this paper cites.
Overview of the trec 2023 deep learning track,
N. Craswell, B. Mitra, E. Yilmaz, H. A. Rahmani, D. Campos, J. Lin, E. M. Voorhees, I. Soboroff, · 2023
Cited alongside, same era.
Llm4eval: Large language model for evaluation in ir,
H. A. Rahmani, C. Siro, M. Aliannejadi, N. Craswell, C. L. A. Clarke, G. Faggioli, B. Mitra, P. Thomas, E. Yilmaz, · 2024
Cited alongside, same era.
Synthetic test collections for retrieval evaluation,
H. A. Rahmani, N. Craswell, E. Yilmaz, B. Mitra, D. Campos, · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…