Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have experienced widespread adoption across scientific and industrial domains due to their versatility and utility for diverse tasks.
1910
Earlier work this paper cites.
1911
Earlier work this paper cites.
W. Reese, “Nginx: the high-performance web server and reverse proxy,” Linux J. , vol. 2008, no. 173, sep 2008
2008
Earlier work this paper cites.
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel et al. , “Scikit-learn: Machine learning in Python,” vol. 12, pp. 2825–2830, 2011
2011
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez et al. , “Attention is all you need,” in Advances in Neural Information Processing Systems , I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett, Eds., vol. 30. Curran Associates, Inc., 2017
2017
Earlier work this paper cites.
D. Buchaca, J. L. Berral, C. Wang, and A. Youssef, “Proactive container auto-scaling for cloud native machine learning services,” in 2020 IEEE 13th International Conference on Cloud Computing (CLOUD) , 2020, pp. 475–479
2020
Earlier work this paper cites.
R. Bommasani, D. A. Hudson, E. Adeli, R. Altman, S. Arora, S. von Arx et al. , “On the opportunities and risks of foundation models,” 2022
2022
Earlier work this paper cites.
P. Liang, R. Bommasani, T. Lee, D. Tsipras, D. Soylu, M. Yasunaga et al. , “Holistic evaluation of language models,” 2022
2022
Earlier work this paper cites.
G.-I. Yu, J. S. Jeong, G.-W. Kim, S. Kim, and B.-G. Chun, “Orca: A distributed serving system for Transformer-Based generative models,” in 16th USENIX Symposium on Operating Systems Design and Implementation (OSDI 22) . Carlsbad, CA: USENIX Association, Jul. 2022, pp. 521–538. [Online]. Available: https://www.usenix.org/conference/osdi22/presentation/yu
2022
Cited alongside, same era.
M. R. A. H. Rony, C. Suess, S. R. Bhat, V. Sudhi, J. Schneider, M. Vogel et al. , “Carexpert: Leveraging large language models for in-car conversational question answering,” 2023
2023
Cited alongside, same era.
N. Gruver, A. Sriram, A. Madotto, A. Wilson, C. L. Zitnick, and Z. Ulissi, “Fine-tuned language models generate stable inorganic materials as text,” in AI for Accelerated Materials Design - NeurIPS 2023 Workshop , 2023. [Online]. Available: https://openreview.net/forum?id=0r5DE2ZSwJ
2023
Cited alongside, same era.
Z. Zheng, X. Ren, F. Xue, Y. Luo, X. Jiang, and Y. You, “Response length perception and sequence scheduling: An llm-empowered llm inference pipeline,” in Advances in Neural Information Processing Systems , vol. 36. Curran Associates, Inc., 2023, pp. 65 517–65 530. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2023/file/ce7ff3405c782f761fac7f849b41ae9a-Paper-Conference.pdf
2023
Later among the works it cites.
OpenAI, J. Achiam, S. Adler, S. Agarwal, L. Ahmad, I. Akkaya et al. , “Gpt-4 technical report,” 2024
2024
Closest in time.
Anthropic, “The claude 3 model family: Opus, sonnet, haiku,” https://www-cdn.anthropic.com/de8ba9b01c9ab7cbabf5c33b80b7bbc618857627/Model_Card_Claude_3.pdf , 2024, accessed: 2024-03-31
2024
Closest in time.
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
S. Yao, D. Yu, J. Zhao, I. Shafran, T. L. Griffiths, Y. Cao et al. , “Tree of thoughts: Deliberate problem solving with large language models,” 2023
2023
Cited alongside, same era.
Q. Wu, G. Bansal, J. Zhang, Y. Wu, B. Li, E. Zhu et al. , “Autogen: Enabling next-gen llm applications via multi-agent conversation framework,” 2023
2023
Cited alongside, same era.
S. Gururangan, M. Li, M. Lewis, W. Shi, T. Althoff, N. A. Smith et al. , “Scaling expert language models with unsupervised domain discovery,” 2023
2023
Cited alongside, same era.
Y. Wang, K. Chen, H. Tan, and K. Guo, “Tabi: An efficient multi-level inference system for large language models,” in Proceedings of the Eighteenth European Conference on Computer Systems , ser. EuroSys ’23. New York, NY, USA: Association for Computing Machinery, 2023, p. 233–248. [Online]. Available: https://doi.org/10.1145/3552326.3587438
2023
Cited alongside, same era.
Cited in the paper.
Cited in the paper.
Cited in the paper.
Cited in the paper.
N. Alshahwan, M. Harman, I. Harper, A. Marginean, S. Sengupta, and E. Wang, “Assured llm-based software engineering,” 2024
2024
Closest in time.
T. Atta-fosu, A. Arunkumar, A. Lokhmotov, A. Nanjappa, H. Itay, M. Szutenberg et al. Llama 2 70b: An MLPerf inference benchmark for large language models. [Online]. Available: https://mlcommons.org/2024/03/mlperf-llama2-70b/
2024
Closest in time.
gRPC Authors, “grpc,” 2023, accessed: 2024-03-29. [Online]. Available: https://grpc.io/
2024
Closest in time.