Fetching the paper…
Reading the bibliography…
The rise of large language models (LLMs) is revolutionizing information retrieval, question answering, summarization, and code generation tasks.
C. E. Shannon, “A mathematical theory of communication,” The Bell System Technical Journal , vol. 27, no. 3, pp. 379–423, 1948
1948
Earlier work this paper cites.
T. Mikolov, K. Chen, G. Corrado, and J. Dean, “Efficient estimation of word representations in vector space,” 2013
2013
Earlier work this paper cites.
U. Alon, M. Zilberstein, O. Levy, and E. Yahav, “code2vec: Learning distributed representations of code,” 2018
2018
Earlier work this paper cites.
S. Takahashi and K. Tanaka-Ishii, “Evaluating Computational Language Models with Scaling Properties of Natural Language,” Computational Linguistics , vol. 45, no. 3, pp. 481–513, 09 2019. [Online]. Available: https://doi.org/10.1162/coli_a_00355
2019
Earlier work this paper cites.
M. Chen, J. Tworek, H. Jun, Q. Yuan, H. P. de Oliveira Pinto, J. Kaplan, H. Edwards, Y. Burda, N. Joseph, G. Brockman, A. Ray, R. Puri, G. Krueger, M. Petrov, H. Khlaaf, G. Sastry, P. Mishkin, B. Chan, S. Gray, N. Ryder, M. Pavlov, A. Power, L. Kaiser, M. Bavarian, C. Winter, P. Tillet, F. P. Such, D. Cummings, M. Plappert, F. Chantzis, E. Barnes, A. Herbert-Voss, W. H. Guss, A. Nichol, A. Paino, N. Tezak, J. Tang, I. Babuschkin, S. Balaji, S. Jain, W. Saunders, C. Hesse, A. N. Carr, J. Leike, J. Achiam, V. Misra, E. Morikawa, A. Radford, M. Knight, M. Brundage, M. Murati, K. Mayer, P. Welinder, B. McGrew, D. Amodei, S. McCandlish, I. Sutskever, and W. Zaremba, “Evaluating large language models trained on code,” 2021
2021
Earlier work this paper cites.
H. Pearce, B. Ahmad, B. Tan, B. Dolan-Gavitt, and R. Karri, “Asleep at the keyboard? assessing the security of github copilot’s code contributions,” 2021
2021
Earlier work this paper cites.
C. Meister and R. Cotterell, “Language model evaluation beyond perplexity,” 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
“Zlib technical details,” Zlib.net , 2022
2022
Earlier work this paper cites.
Google, “Embeddings,” 2022. [Online]. Available: {https://developers.google.com/machine-learning/crash-course/embeddings/video-lecture}
2022
Earlier work this paper cites.
R. S, A. S. Bharadwaj, D. S K, M. S. Khadabadi, and A. Jayaprakash, “Digital implementation of the softmax activation function and the inverse softmax function,” in 2022 4th International Conference on Circuits, Control, Communication and Computing (I4C) , 2022, pp. 64–67
2022
Earlier work this paper cites.
Y. Yang, S. Mandt, and L. Theis, “An introduction to neural data compression,” 2022
2022
Earlier work this paper cites.
M. Alwani, Y. Wang, and V. Madhavan, “Decore: Deep compression with reinforcement learning,” 2022
2022
Cited alongside, same era.
A. M. Dakhel, V. Majdinasab, A. Nikanjam, F. Khomh, M. C. Desmarais, Z. Ming, and Jiang, “Github copilot ai pair programmer: Asset or liability?” 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
P. Liang, R. Bommasani, T. Lee, D. Tsipras, D. Soylu, M. Yasunaga, Y. Zhang, D. Narayanan, Y. Wu, A. Kumar, B. Newman, B. Yuan, B. Yan, C. Zhang, C. Cosgrove, C. D. Manning, C. Ré, D. Acosta-Navas, D. A. Hudson, E. Zelikman, E. Durmus, F. Ladhak, F. Rong, H. Ren, H. Yao, J. Wang, K. Santhanam, L. Orr, L. Zheng, M. Yuksekgonul, M. Suzgun, N. Kim, N. Guha, N. Chatterji, O. Khattab, P. Henderson, Q. Huang, R. Chi, S. M. Xie, S. Santurkar, S. Ganguli, T. Hashimoto, T. Icard, T. Zhang, V. Chaudhary, W. Wang, X. Li, Y. Mai, Y. Zhang, and Y. Koreeda, “Holistic evaluation of language models,” 2022
——, “Models overview,” 2023. [Online]. Available: {https://platform.openai.com/docs/models/gpt-3-5}
2023
Closest in time.
Adobe, “Lossy vs lossless compression differences and when to use.” 2023. [Online]. Available: {https://www.adobe.com/uk/creativecloud/photography/discover/lossy-vs-lossless.html}
2023
Closest in time.
P. S. Foundation, “zlib — compression compatible with gzip,” 2023. [Online]. Available: {https://docs.python.org/3/library/zlib.html}
2023
Closest in time.
Wikipedia, “Levenshtein distance,” 2023. [Online]. Available: {https://en.wikipedia.org/wiki/Levenshtein_distance}
2023
Closest in time.
OpenAI, “Embeddings,” 2023. [Online]. Available: {https://platform.openai.com/docs/guides/embeddings}
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
OpenAI, “Introducing chatgpt,” 2023. [Online]. Available: {https://openai.com/blog/chatgpt}
2023
Cited alongside, same era.
Google, “Bard,” 2023. [Online]. Available: {https://bard.google.com/}
2023
Cited alongside, same era.
Anthropic, “Introducing claude,” 2023. [Online]. Available: {https://www.anthropic.com/index/introducing-claude}
2023
Cited alongside, same era.
A. AWS, “Amazon titan,” 2023. [Online]. Available: {https://aws.amazon.com/bedrock/titan/}
2023
Cited alongside, same era.
AI21, “Announcing ai21 studio and jurassic-1 language models,” 2023. [Online]. Available: {https://www.ai21.com/blog/announcing-ai21-studio-and-jurassic-1}
2023
Cited alongside, same era.
J. White, Q. Fu, S. Hays, M. Sandborn, C. Olea, H. Gilbert, A. Elnashar, J. Spencer-Smith, and D. C. Schmidt, “A prompt pattern catalog to enhance prompt engineering with chatgpt,” 2023
2023
Cited alongside, same era.
OpenAI, “What are tokens and how to count them?” 2023. [Online]. Available: {https://help.openai.com/en/articles/4936856-what-are-tokens-and-how-to-count-them}
2023
Cited alongside, same era.
“Github copilot · your ai pair programmer.” [Online]. Available: https://github.com/features/copilot
Cited in the paper.
Wikipedia, “Cosine similarity,” 2023. [Online]. Available: {https://en.wikipedia.org/wiki/Cosine_similarity}
2023
Closest in time.
2023
Closest in time.
G. Sandoval, H. Pearce, T. Nys, R. Karri, S. Garg, and B. Dolan-Gavitt, “Lost at c: A user study on the security implications of large language model code assistants,” 2023
2023
Closest in time.
S. Bubeck, V. Chandrasekaran, R. Eldan, J. Gehrke, E. Horvitz, E. Kamar, P. Lee, Y. T. Lee, Y. Li, S. Lundberg, H. Nori, H. Palangi, M. T. Ribeiro, and Y. Zhang, “Sparks of artificial general intelligence: Early experiments with gpt-4,” 2023
2023
Closest in time.
2023
Closest in time.
P. Liu, W. Yuan, J. Fu, Z. Jiang, H. Hayashi, and G. Neubig, “Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing,” ACM Computing Surveys , vol. 55, no. 9, pp. 1–35, 2023
2023
Closest in time.