Fetching the paper…
Reading the bibliography…
Neural Information Retrieval has advanced rapidly in high-resource languages, but progress in lower-resource ones such as Japanese has been hindered by data scarcity, among other challenges.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 1901
Earlier work this paper cites.
Retrieval techniques
Belkin, N., and Croft, W · 1987
Earlier work this paper cites.
New stochastic approximation type procedures
Polyak, B · 1990
Earlier work this paper cites.
Acceleration of stochastic approximation by averaging
Polyak, B. T., and Juditsky, A. B · 1992
Earlier work this paper cites.
Stochastic approximation via averaging: the polyak’s approach revisited
Yin, G · 1992
Earlier work this paper cites.
Okapi at trec-3
Robertson, S. E., Walker, S., Jones, S., Hancock-Beaulieu, M. M., Gatford, M., et al · 1995
Earlier work this paper cites.
Modern information retrieval
Baeza-Yates, R., Ribeiro-Neto, B., et al · 1999
Earlier work this paper cites.
Information retrieval on the web
Kobayashi, M., and Takeda, K · 2000
Earlier work this paper cites.
Introduction to information retrieval
Manning, C. D · 2008
Earlier work this paper cites.
Improvements to bm25 and language models examined
Trotman, A., Puurula, A., and Burgess, B · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network, 2015
Hinton, G., Vinyals, O., and Dean, J · 2015
Earlier work this paper cites.
MS MARCO: A human generated machine reading comprehension dataset
Nguyen, T., Rosenberg, M., Song, X., Gao, J., Tiwary, S., Majumder, R., and Deng, L · 2016
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text
Rajpurkar, P., Zhang, J., Lopyrev, K., and Liang, P · 2016
Earlier work this paper cites.
Deep learning scaling is predictable, empirically
Hestness, J., Narang, S., Ardalani, N., Diamos, G., Jun, H., Kianinejad, H., Patwary, M. M. A., Yang, Y., and Zhou, Y · 2017
Earlier work this paper cites.
Lstm vs. bm25 for open-domain qa: A hands-on comparison of effectiveness and efficiency
Kato, S., Togashi, R., Maeda, H., Fujita, S., and Sakai, T · 2017
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
Kirkpatrick, J., Pascanu, R., Rabinowitz, N., Veness, J., Desjardins, G., Rusu, A. A., Milan, K., Quan, J., Ramalho, T., Grabska-Barwinska, A., et al · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Loshchilov, I., and Hutter, F · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Earlier work this paper cites.
The marginal value of adaptive gradient methods in machine learning
Wilson, A. C., Roelofs, R., Stern, M., Srebro, N., and Recht, B · 2017
Earlier work this paper cites.
An introduction to neural information retrieval
Mitra, B., Craswell, N., et al · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2019
Earlier work this paper cites.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Sanh, V., Debut, L., Chaumond, J., and Wolf, T · 2019
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Conneau, A., Khandelwal, K., Goyal, N., Chaudhary, V., Wenzek, G., Guzmán, F., Grave, É., Ott, M., Zettlemoyer, L., and Stoyanov, V · 2020
Earlier work this paper cites.
Overview of the trec 2019 deep learning track
Craswell, N., Mitra, B., Yilmaz, E., Campos, D., and Voorhees, E. M · 2020
Earlier work this paper cites.
Improving efficient neural ranking models with cross-architecture knowledge distillation
Hofstätter, S., Althammer, S., Schröder, M., Sertkan, M., and Hanbury, A · 2020
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering
Karpukhin, V., Oguz, B., Min, S., Lewis, P., Wu, L., Edunov, S., Chen, D., and Yih, W.-t · 2020
Earlier work this paper cites.
Colbert: Efficient and effective passage search via contextualized late interaction over bert
Khattab, O., and Zaharia, M · 2020
Earlier work this paper cites.
Document ranking with a pretrained sequence-to-sequence model
Nogueira, R., Jiang, Z., and Lin, J · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P. J · 2020
Cited alongside, same era.
Minilm: Deep self-attention distillation for task-agnostic compression of pre-trained transformers
Wang, W., Wei, F., Dong, L., Bao, H., Yang, N., and Zhou, M · 2020
Cited alongside, same era.
LUKE: Deep contextualized entity representations with entity-aware self-attention
Yamada, I., Asai, A., Shindo, H., Takeda, H., and Matsumoto, Y · 2020
Cited alongside, same era.
MMARCO: A multilingual version of the MS MMARCO passage ranking dataset
Bonifacio, L., Jeronymo, V., Abonizio, H. Q., Campiotti, I., Fadaee, M., Lotufo, R., and Nogueira, R · 2021
Cited alongside, same era.
Splade: Sparse lexical and expansion model for first stage ranking
Formal, T., Piwowarski, B., and Clinchant, S · 2021
Cited alongside, same era.
Achiam, J., Adler, S., Agarwal, S., Ahmad, L., Akkaya, I., Aleman, F. L., Almeida, D., Altenschmidt, J., Altman, S., Anadkat, S., et al · 2023
Later among the works it cites.
Clavié, B · 2023
Later among the works it cites.
Simple yet effective neural ranking and reranking baselines for cross-lingual information retrieval
Lin, J., Alfonso-Hermelo, D., Jeronymo, V., Kamalloo, E., Lassance, C., Nogueira, R., Ogundepo, O., Rezagholizadeh, M., Thakur, N., Yang, J.-H., et al · 2023
Later among the works it cites.
Udapdr: Unsupervised domain adaptation via llm prompting and distillation of rerankers
Saad-Falcon, J., Khattab, O., Santhanam, K., Florian, R., Franz, M., Roukos, S., Sil, A., Sultan, M., and Potts, C · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A white box analysis of colbert
Formal, T., Piwowarski, B., and Clinchant, S · 2021
Cited alongside, same era.
Distilling knowledge from reader to retriever for question answering
Izacard, G., and Grave, E · 2021
Cited alongside, same era.
Comparing kullback-leibler divergence and mean squared error loss in knowledge distillation
Kim, T., Oh, J., Kim, N., Cho, S., and Yun, S.-Y · 2021
Cited alongside, same era.
In-batch negatives for knowledge distillation with tightly-coupled teachers for dense retrieval
Lin, S.-C., Yang, J.-H., and Lin, J · 2021
Cited alongside, same era.
Rocketqa: An optimized training approach to dense passage retrieval for open-domain question answering
Qu, Y., Ding, Y., Liu, J., Liu, K., Ren, R., Zhao, W. X., Dong, D., Wu, H., and Wang, H · 2021
Cited alongside, same era.
Rocketqav2: A joint training method for dense passage retrieval and passage re-ranking
Ren, R., Qu, Y., Liu, J., Zhao, W. X., She, Q., Wu, H., Wang, H., and Wen, J.-R · 2021
Cited alongside, same era.
BEIR: A heterogenous benchmark for zero-shot evaluation of information retrieval models
Thakur, N., Reimers, N., Rücklé, A., Srivastava, A., and Gurevych, I · 2021
Cited alongside, same era.
Touvron, H., Martin, L., Stone, K., Albert, P., Almahairi, A., Babaei, Y., Bashlykov, N., Batra, S., Bhargava, P., Bhosale, S., et al · 2023
Later among the works it cites.
Japanese SimCSE technical report
Tsukagoshi, H., Sasano, R., and Takeda, K · 2023
Later among the works it cites.
C-pack: Packaged resources to advance general chinese embedding
Xiao, S., Liu, Z., Zhang, P., and Muennighof, N · 2023
Later among the works it cites.
C-pack: Packaged resources to advance general chinese embedding
Xiao, S., Liu, Z., Zhang, P., and Muennighof, N · 2023
Later among the works it cites.
MIRACL: A multilingual retrieval dataset covering 18 diverse languages
Zhang, X., Thakur, N., Ogundepo, O., Kamalloo, E., Alfonso-Hermelo, D., Li, X., Liu, Q., Rezagholizadeh, M., and Lin, J · 2023
Later among the works it cites.
Evolutionary optimization of model merging recipes
Akiba, T., Shing, M., Tang, Y., Sun, Q., and Ha, D · 2024
Closest in time.
Chen, J., Xiao, S., Zhang, P., Luo, K., Lian, D., and Liu, Z · 2024
Closest in time.
Defazio, A., Xingyu, Yang, Mehta, H., Mishchenko, K., Khaled, A., and Cutkosky, A · 2024
Closest in time.
Beneath the [mask]: An analysis of structural query tokens in colbert
Giacalone, B., Paiement, G., Tucker, Q., and Zanibbi, R · 2024
Closest in time.
Minicpm: Unveiling the potential of small language models with scalable training strategies
Hu, S., Tu, Y., Han, X., He, C., Cui, G., Long, X., Zheng, Z., Fang, Y., Huang, Y., Zhao, W., et al · 2024
Closest in time.
Splade-v3: New baselines for splade
Lassance, C., Déjean, H., Formal, T., and Clinchant, S · 2024
Closest in time.
Rethinking the role of token retrieval in multi-vector retrieval
Lee, J., Dai, Z., Duddu, S. M. K., Lei, T., Naim, I., Chang, M.-W., and Zhao, V · 2024
Closest in time.
The llama 3 herd of models
LlamaTeam · 2024
Closest in time.
Decouvrir: A benchmark for evaluating the robustness of information retrieval models in french, 2024
Louis, A · 2024
Closest in time.
Louis, A., Saxena, V., van Dijck, G., and Spanakis, G · 2024
Closest in time.
A reproducibility study of plaid
MacAvaney, S., and Tonellotto, N · 2024
Closest in time.
Arctic-embed: Scalable, efficient, and accurate text embedding models
Merrick, L., Xu, D., Nuti, G., and Campos, D · 2024
Closest in time.
Jacwir: Japanese casual web ir
Tateno, Y · 2024
Closest in time.
Jqara: Japanese question answering with retrieval augmentation
Tateno, Y · 2024
Closest in time.
Seaeval for multilingual foundation models: From cross-lingual alignment to cultural reasoning
Wang, B., Liu, Z., Huang, X., Jiao, F., Ding, Y., Aw, A., and Chen, N · 2024
Closest in time.
Multilingual e5 text embeddings: A technical report
Wang, L., Yang, N., Huang, X., Yang, L., Majumder, R., and Wei, F · 2024
Closest in time.
Translate-distill: Learning cross-language dense retrieval by translation and distillation
Yang, E., Lawrie, D., Mayfield, J., Oard, D. W., and Miller, S · 2024
Closest in time.
日本語 Reranker 作成のテクニカルレポート
舘野 祐一 · 2024
Closest in time.