Fetching the paper…
Reading the bibliography…
Retrieval Augmented Generation (RAG) has become one of the most popular paradigms for enabling LLMs to access external data, and also as a mechanism for grounding to mitigate against hallucinations.
1905
Earlier work this paper cites.
1910
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th annual meeting on association for computational linguistics . Association for Computational Linguistics, 2002, pp. 311–318
2002
Earlier work this paper cites.
2003
Earlier work this paper cites.
C.-Y. Lin, “Rouge: a package for automatic evaluation of summaries,” 2004
2004
Earlier work this paper cites.
G. V. Cormack, C. L. A. Clarke, and S. Büttcher, “Reciprocal rank fusion outperforms condorcet and individual rank learning methods,” Proceedings of the 32nd international ACM SIGIR conference on Research and development in information retrieval , 2009. [Online]. Available: https://api.semanticscholar.org/CorpusID:12408211
2009
Earlier work this paper cites.
2010
Earlier work this paper cites.
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang, “Squad: 100,000+ questions for machine comprehension of text,” 2016
2016
Earlier work this paper cites.
M. Henderson, R. Al-Rfou, B. Strope, Y. hsuan Sung, L. Lukacs, R. Guo, S. Kumar, B. Miklos, and R. Kurzweil, “Efficient natural language response suggestion for smart reply,” 2017
2017
Earlier work this paper cites.
A. Radford and K. Narasimhan, “Improving language understanding by generative pre-training,” 2018. [Online]. Available: https://api.semanticscholar.org/CorpusID:49313245
2018
Earlier work this paper cites.
B. Mitra and N. Craswell, “An introduction to neural information retrieval,” Foundations and Trends® in Information Retrieval , vol. 13, no. 1, pp. 1–126, 2018. [Online]. Available: http://dx.doi.org/10.1561/1500000061
2018
Earlier work this paper cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language models are unsupervised multitask learners,” 2019. [Online]. Available: https://api.semanticscholar.org/CorpusID:160025533
2019
Earlier work this paper cites.
J. Guo, Y. Fan, L. Pang, L. Yang, Q. Ai, H. Zamani, C. Wu, W. B. Croft, and X. Cheng, “A deep look into neural ranking models for information retrieval,” 2019
2019
Cited alongside, same era.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” 2020
2020
Cited alongside, same era.
D. Brown, “Rank-BM25: A Collection of BM25 Algorithms in Python,” 2020. [Online]. Available: https://doi.org/10.5281/zenodo.4520057
2020
Cited alongside, same era.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W. tau Yih, T. Rocktäschel, S. Riedel, and D. Kiela, “Retrieval-augmented generation for knowledge-intensive nlp tasks,” 2021
Y. Zhang, Y. Li, L. Cui, D. Cai, L. Liu, T. Fu, X. Huang, E. Zhao, Y. Zhang, Y. Chen, L. Wang, A. T. Luu, W. Bi, F. Shi, and S. Shi, “Siren’s song in the ai ocean: A survey on hallucination in large language models,” 2023
2023
Later among the works it cites.
Y. Gao, Y. Xiong, X. Gao, K. Jia, J. Pan, Y. Bi, Y. Dai, J. Sun, and H. Wang, “Retrieval-augmented generation for large language models: A survey,” 2023
2023
Later among the works it cites.
Y. Liu, D. Iter, Y. Xu, S. Wang, R. Xu, and C. Zhu, “G-eval: Nlg evaluation using gpt-4 with better human alignment,” 2023
2023
Later among the works it cites.
E. Horvitz. (2023) The power of prompt. [Online]. Available: https://www.microsoft.com/en-us/research/blog/the-power-of-prompting/
2023
Later among the works it cites.
S. Chen, S. Wong, L. Chen, and Y. Tian, “Extending context window of large language models via positional interpolation,” 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
G. M. Rosa, R. C. Rodrigues, R. Lotufo, and R. Nogueira, “Yes, bm25 is a strong baseline for legal case retrieval,” 2021
2021
Cited alongside, same era.
J. Lin, X. Ma, S.-C. Lin, J.-H. Yang, R. Pradeep, and R. Nogueira, “Pyserini: An easy-to-use python toolkit to support replicable ir research with sparse and dense representations,” 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
J. Pereira, R. Fidalgo, R. Lotufo, and R. Nogueira, “Visconde: Multi-document qa with gpt-3 and neural reranking,” 2022
2022
Cited alongside, same era.
O. Press, N. A. Smith, and M. Lewis, “Train short, test long: Attention with linear biases enables input length extrapolation,” 2022
2022
Cited alongside, same era.
J. Liu, “Llamaindex,” 2022. [Online]. Available: https://github.com/jerryjliu/llama_index
2022
Cited alongside, same era.
W. X. Zhao, J. Liu, R. Ren, and J.-R. Wen, “Dense text retrieval based on pretrained language models: A survey,” 2022
2022
Cited alongside, same era.
2023
Later among the works it cites.
P. Xu, W. Ping, X. Wu, L. McAfee, C. Zhu, Z. Liu, S. Subramanian, E. Bakhturina, M. Shoeybi, and B. Catanzaro, “Retrieval meets long context large language models,” 2023
2023
Later among the works it cites.
N. F. Liu, K. Lin, J. Hewitt, A. Paranjape, M. Bevilacqua, F. Petroni, and P. Liang, “Lost in the middle: How language models use long contexts,” 2023
2023
Later among the works it cites.
G. Kamradt, “Pressure testing claude-2.1 200k via needle-in-a-haystack,” 2023. [Online]. Available: https://github.com/gkamradt/LLMTest_NeedleInAHaystack
2023
Later among the works it cites.
G. Mohandas and P. Moritz. (2023) Building rag-based llm applications for production. [Online]. Available: https://www.anyscale.com/blog/a-comprehensive-guide-for-building-rag-based-llm-applications-part-1
2023
Later among the works it cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” 2023
2023
Later among the works it cites.
G. Alperovich, “The secret sauce behind 100k context window in llms: all tricks in one place,” https://blog.gopenai.com/how-to-speed-up-llms-and-use-100k-context-window-all-tricks-in-one-place-ffd40577b4c , 2023/May, accessed: 01/02/2024
2024
Closest in time.