Fetching the paper…
Reading the bibliography…
The robustness of large language models (LLMs) becomes increasingly important as their use rapidly grows in a wide range of domains.
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
Y. A. Malkov and D. A. Yashunin, “Efficient and robust approximate nearest neighbor search using hierarchical navigable small world graphs,” IEEE transactions on pattern analysis and machine intelligence , vol. 42, no. 4, pp. 824–836, 2018
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel et al. , “Retrieval-augmented generation for knowledge-intensive nlp tasks,” Advances in Neural Information Processing Systems , vol. 33, pp. 9459–9474, 2020
2020
Earlier work this paper cites.
J. X. Morris, E. Lifland, J. Y. Yoo, and Y. Qi, “Textattack: A framework for adversarial attacks in natural language processing,” Proceedings of the 2020 EMNLP, Arvix , 2020
2020
Earlier work this paper cites.
D. Jin, Z. Jin, J. T. Zhou, and P. Szolovits, “Is bert really robust? a strong baseline for natural language attack on text classification and entailment,” in Proceedings of the AAAI conference on artificial intelligence , vol. 34, no. 05, 2020, pp. 8018–8025
2020
Earlier work this paper cites.
A. Lazaridou, A. Kuncoro, E. Gribovskaya, D. Agrawal, A. Liska, T. Terzi, M. Gimenez, C. de Masson d’Autume, T. Kocisky, S. Ruder et al. , “Mind the gap: Assessing temporal generalization in neural language models,” Advances in Neural Information Processing Systems , vol. 34, pp. 29 348–29 363, 2021
2021
Earlier work this paper cites.
B. Wang and A. Komatsuzaki, “Gpt-j-6b: A 6 billion parameter autoregressive language model.” https://github.com/kingoflolz/mesh-transformer-jax , May 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
S. Borgeaud, A. Mensch, J. Hoffmann, T. Cai, E. Rutherford, K. Millican, G. B. Van Den Driessche, J.-B. Lespiau, B. Damoc, A. Clark et al. , “Improving language models by retrieving from trillions of tokens,” in International conference on machine learning . PMLR, 2022, pp. 2206–2240
2022
Earlier work this paper cites.
K. Meng, D. Bau, A. Andonian, and Y. Belinkov, “Locating and editing factual associations in gpt,” Advances in Neural Information Processing Systems , vol. 35, pp. 17 359–17 372, 2022
2022
Cited alongside, same era.
D. Ganguli, D. Hernandez, L. Lovitt, A. Askell, Y. Bai, A. Chen, T. Conerly, N. Dassarma, D. Drain, N. Elhage et al. , “Predictability and surprise in large generative models,” in Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency , 2022, pp. 1747–1764
2022
Cited alongside, same era.
E. Mitchell, C. Lin, A. Bosselut, C. D. Manning, and C. Finn, “Memory-based model editing at scale,” in International Conference on Machine Learning . PMLR, 2022, pp. 15 817–15 831
2022
Cited alongside, same era.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou et al. , “Chain-of-thought prompting elicits reasoning in large language models,” Advances in Neural Information Processing Systems , vol. 35, pp. 24 824–24 837, 2022
2022
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Cited alongside, same era.
A. Mallen, A. Asai, V. Zhong, R. Das, D. Khashabi, and H. Hajishirzi, “When not to trust language models: Investigating effectiveness of parametric and non-parametric memories,” in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2023, pp. 9802–9822
2023
Cited alongside, same era.
N. Kandpal, H. Deng, A. Roberts, E. Wallace, and C. Raffel, “Large language models struggle to learn long-tail knowledge,” in International Conference on Machine Learning . PMLR, 2023, pp. 15 696–15 707
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
B. Liu, B. Xiao, X. Jiang, S. Cen, X. He, W. Dou et al. , “Adversarial attacks on large language model-based system and mitigating strategies: A case study on chatgpt,” Security and Communication Networks , vol. 2023, 2023
2023
Later among the works it cites.
F. Shi, X. Chen, K. Misra, N. Scales, D. Dohan, E. H. Chi, N. Schärli, and D. Zhou, “Large language models can be easily distracted by irrelevant context,” in International Conference on Machine Learning . PMLR, 2023, pp. 31 210–31 227
2023
Later among the works it cites.
R. Pradeep, K. Hui, J. Gupta, A. Lelkes, H. Zhuang, J. Lin, D. Metzler, and V. Tran, “How does generative retrieval scale to millions of passages?” in Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , H. Bouamor, J. Pino, and K. Bali, Eds. Singapore: Association for Computational Linguistics, Dec. 2023, pp. 1305–1321. [Online]. Available: https://aclanthology.org/2023.emnlp-main.83
2023
Later among the works it cites.
IMDb, “Imdb datasets,” 2023, accessed: 2023-11-20. [Online]. Available: https://developer.imdb.com/non-commercial-datasets/
2023
Later among the works it cites.
Wikipedia Contributors, “Wikipedia, the free encyclopedia,” 2023, accessed: 2023-11-20. [Online]. Available: https://www.wikipedia.org
2023
Later among the works it cites.
Opendatasoft, “nobel prize,” https://public.opendatasoft.com/explore/dataset/nobel-prize-laureates , 2023, accessed: 2023-11-20
2023
Later among the works it cites.
Rui Meng, Ye Liu, Shafiq Rayhan Joty, Caiming Xiong, Yingbo Zhou, Semih Yavuz, “Sfr-embedding-mistral:enhance text retrieval with transfer learning,” Salesforce AI Research Blog, 2024. [Online]. Available: https://blog.salesforceairesearch.com/sfr-embedded-mistral/
2024
Closest in time.