Fetching the paper…
Reading the bibliography…
The emergence of Large Language Models (LLMs) has reshaped agent systems.
W. Van Melle, “Mycin: a knowledge-based consultation program for infectious disease diagnosis,” International journal of man-machine studies , vol. 10, no. 3, pp. 313–322, 1978
1978
Earlier work this paper cites.
B. G. Buchanan and E. A. Feigenbaum, “Dendral and meta-dendral: Their applications dimension,” in Readings in artificial intelligence . Elsevier, 1981, pp. 313–322
1981
Earlier work this paper cites.
M. Wooldridge and N. R. Jennings, “Intelligent agents: Theory and practice,” The knowledge engineering review , vol. 10, no. 2, pp. 115–152, 1995
1995
Earlier work this paper cites.
E. Oliveira, K. Fischer, and O. Stepankova, “Multi-agent systems: which research for which applications,” Robotics and Autonomous Systems , vol. 27, no. 1-2, pp. 91–106, 1999
1999
Earlier work this paper cites.
R. S. Sutton, D. McAllester, S. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” Advances in neural information processing systems , vol. 12, 1999
1999
Earlier work this paper cites.
T. H. Davenport and J. G. Harris, “Automated decision making comes of age,” MIT Sloan Management Review , vol. 46, no. 4, p. 83, 2005
2005
Earlier work this paper cites.
H. Van Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double q-learning,” in Proceedings of the AAAI conference on artificial intelligence , vol. 30, no. 1, 2016
2016
Earlier work this paper cites.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 26–38, 2017
2017
Earlier work this paper cites.
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton et al. , “Mastering the game of go without human knowledge,” nature , vol. 550, no. 7676, pp. 354–359, 2017
2017
Earlier work this paper cites.
A. Dorri, S. S. Kanhere, and R. Jurdak, “Multi-agent systems: A survey,” Ieee Access , vol. 6, pp. 28 573–28 593, 2018
2018
Earlier work this paper cites.
P. Hernandez-Leal, B. Kartal, and M. E. Taylor, “A survey and critique of multiagent deep reinforcement learning,” Autonomous Agents and Multi-Agent Systems , vol. 33, no. 6, pp. 750–797, 2019
2019
Earlier work this paper cites.
H. Jang, J. Kim, J.-E. Jo, J. Lee, and J. Kim, “Mnnfast: A fast and scalable system architecture for memory-augmented neural networks,” in Proceedings of the 46th International Symposium on Computer Architecture , 2019, pp. 250–263
2019
Earlier work this paper cites.
T. T. Nguyen, N. D. Nguyen, and S. Nahavandi, “Deep reinforcement learning for multiagent systems: A review of challenges, solutions, and applications,” IEEE transactions on cybernetics , vol. 50, no. 9, pp. 3826–3839, 2020
2020
Earlier work this paper cites.
J. Clifton and E. Laber, “Q-learning: Theory and applications,” Annual Review of Statistics and Its Application , vol. 7, no. 1, pp. 279–301, 2020
2020
Earlier work this paper cites.
A. Kumar, A. Zhou, G. Tucker, and S. Levine, “Conservative q-learning for offline reinforcement learning,” Advances in neural information processing systems , vol. 33, pp. 1179–1191, 2020
2020
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel et al. , “Retrieval-augmented generation for knowledge-intensive nlp tasks,” Advances in neural information processing systems , vol. 33, pp. 9459–9474, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
S. Lu, M. Wang, S. Liang, J. Lin, and Z. Wang, “Hardware accelerator for multi-head attention and position-wise feed-forward in the transformer,” in 2020 IEEE 33rd International System-on-Chip Conference (SOCC) . IEEE, 2020, pp. 84–89
2020
Earlier work this paper cites.
S. Gronauer and K. Diepold, “Multi-agent deep reinforcement learning: a survey,” Artificial Intelligence Review , vol. 55, no. 2, pp. 895–943, 2022
2022
Earlier work this paper cites.
A. M. Hafiz, “A survey of deep q-networks used for reinforcement learning: state of the art,” Intelligent Communication Technologies and Virtual Mobile Networks: Proceedings of ICICV 2022 , pp. 393–402, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Z. Yao, R. Yazdani Aminabadi, M. Zhang, X. Wu, C. Li, and Y. He, “Zeroquant: Efficient and affordable post-training quantization for large-scale transformers,” Advances in Neural Information Processing Systems , vol. 35, pp. 27 168–27 183, 2022
2022
Earlier work this paper cites.
R. Behnia, M. R. Ebrahimi, J. Pacheco, and B. Padmanabhan, “Ew-tune: A framework for privately fine-tuning large language models with differential privacy,” in 2022 IEEE International Conference on Data Mining Workshops (ICDMW) . IEEE, 2022, pp. 560–566
2022
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
S. Dikshit, A. Atiq, M. Shahid, V. Dwivedi, and A. Thusu, “The use of artificial intelligence to optimize the routing of vehicles and reduce traffic congestion in urban areas,” EAI Endorsed Transactions on Energy Web , vol. 10, pp. 1–13, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Z. Ji, N. Lee, R. Frieske, T. Yu, D. Su, Y. Xu, E. Ishii, Y. J. Bang, A. Madotto, and P. Fung, “Survey of hallucination in natural language generation,” ACM computing surveys , vol. 55, no. 12, pp. 1–38, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
V. Kumar, P. Srivastava, A. Dwivedi, I. Budhiraja, D. Ghosh, V. Goyal, and R. Arora, “Large-language-models (llm)-based ai chatbots: Architecture, in-depth analysis and their performance evaluation,” in International Conference on Recent Trends in Image Processing and Pattern Recognition . Springer, 2023, pp. 237–249
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
G. Xiao, J. Lin, M. Seznec, H. Wu, J. Demouth, and S. Han, “Smoothquant: Accurate and efficient post-training quantization for large language models,” in International Conference on Machine Learning . PMLR, 2023, pp. 38 087–38 099
2023
Earlier work this paper cites.
X. Ma, G. Fang, and X. Wang, “Llm-pruner: On the structural pruning of large language models,” Advances in neural information processing systems , vol. 36, pp. 21 702–21 720, 2023
2023
Earlier work this paper cites.
E. Frantar and D. Alistarh, “Sparsegpt: Massive language models can be accurately pruned in one-shot,” in International Conference on Machine Learning . PMLR, 2023, pp. 10 323–10 337
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Bai, H. Zhou, K. Zhao, J. Chen, J. Yu, and K. Wang, “Transformer-opu: An fpga-based overlay processor for transformer networks,” in 2023 IEEE 31st Annual International Symposium on Field-Programmable Custom Computing Machines (FCCM) . IEEE, 2023, pp. 221–221
2023
Cited alongside, same era.
Y. Sheng, L. Zheng, B. Yuan, Z. Li, M. Ryabinin, B. Chen, P. Liang, C. Ré, I. Stoica, and C. Zhang, “Flexgen: High-throughput generative inference of large language models with a single gpu,” in International Conference on Machine Learning . PMLR, 2023, pp. 31 094–31 116
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
K. Yang, Y. Chu, T. Darwin, A. Han, H. Li, H. Wen, Y. Copur-Gencturk, J. Tang, and H. Liu, “Content knowledge identification with multi-agent large language models (llms),” in International Conference on Artificial Intelligence in Education . Springer, 2024, pp. 284–292
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
W. Kwon, Z. Li, S. Zhuang, Y. Sheng, L. Zheng, C. H. Yu, J. Gonzalez, H. Zhang, and I. Stoica, “Efficient memory management for large language model serving with pagedattention,” in Proceedings of the 29th Symposium on Operating Systems Principles , 2023, pp. 611–626
2023
Cited alongside, same era.
H. Yang, M. Li, H. Zhou, Y. Xiao, Q. Fang, and R. Zhang, “One llm is not enough: Harnessing the power of ensemble learning for medical question answering,” medRxiv , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
G. Mialon, C. Fourrier, T. Wolf, Y. LeCun, and T. Scialom, “Gaia: a benchmark for general ai assistants,” in The Twelfth International Conference on Learning Representations , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
A. M. Bran, S. Cox, O. Schilter, C. Baldassari, A. D. White, and P. Schwaller, “Augmenting large language models with chemistry tools,” Nature Machine Intelligence , vol. 6, no. 5, pp. 525–535, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
J. Lin, J. Tang, H. Tang, S. Yang, W.-M. Chen, W.-C. Wang, G. Xiao, X. Dang, C. Gan, and S. Han, “Awq: Activation-aware weight quantization for on-device llm compression and acceleration,” Proceedings of Machine Learning and Systems , vol. 6, pp. 87–100, 2024
2024
Later among the works it cites.
S. Zeng, J. Liu, G. Dai, X. Yang, T. Fu, H. Wang, W. Ma, H. Sun, S. Li, Z. Huang et al. , “Flightllm: Efficient large language model inference with a complete mapping flow on fpgas,” in Proceedings of the 2024 ACM/SIGDA International Symposium on Field Programmable Gate Arrays , 2024, pp. 223–234
2024
Later among the works it cites.
K. Alizadeh, S. I. Mirzadeh, D. Belenko, S. Khatamifard, M. Cho, C. C. Del Mundo, M. Rastegari, and M. Farajtabar, “Llm in a flash: Efficient large language model inference with limited memory,” in Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2024, pp. 12 562–12 584
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Z. Shen, “Llm with tools: A survey,” arXiv preprint arXiv:2409.18807 , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
T. Singh, H. Aditya, V. K. Madisetti, and A. Bahga, “Whispered tuning: Data privacy preservation in fine-tuning llms through differential privacy,” Journal of Software Engineering and Applications , vol. 17, no. 1, pp. 1–22, 2024
2024
Later among the works it cites.
M. Franco, O. Gaggi, and C. E. Palazzi, “Integrating content moderation systems with large language models,” ACM Transactions on the Web , 2024
2024
Later among the works it cites.
U. Kulsum, H. Zhu, B. Xu, and M. d’Amorim, “A case study of llm for automated vulnerability repair: Assessing impact of reasoning and patch validation feedback,” in Proceedings of the 1st ACM International Conference on AI-Powered Software , 2024, pp. 103–111
2024
Later among the works it cites.
Z. Keskin, D. Joosten, N. Klasen, M. Huber, C. Liu, B. Drescher, and R. H. Schmitt, “Llm-enhanced human-machine interaction for adaptive decision making in dynamic manufacturing process environments,” IEEE Access , 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
A. Singh, A. Ehtesham, S. Kumar, and T. T. Khoei, “A survey of the model context protocol (mcp): Standardizing context to enhance large language models (llms),” 2025
2025
Closest in time.
F. Sufi, “Just-in-time news: An ai chatbot for the modern information age.” AI , vol. 6, no. 2, 2025
2025
Closest in time.
W. Kasri, Y. Himeur, H. A. Alkhazaleh, S. Tarapiah, S. Atalla, W. Mansoor, and H. Al-Ahmad, “From vulnerability to defense: The role of large language models in enhancing cybersecurity,” Computation , vol. 13, no. 2, p. 30, 2025
2025
Closest in time.
2025
Closest in time.
S. Sharma, P. Mittal, M. Kumar, and V. Bhardwaj, “The role of large language models in personalized learning: a systematic review of educational impact,” Discover Sustainability , vol. 6, no. 1, pp. 1–24, 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
L. Huang, W. Yu, W. Ma, W. Zhong, Z. Feng, H. Wang, Q. Chen, W. Peng, X. Feng, B. Qin et al. , “A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions,” ACM Transactions on Information Systems , vol. 43, no. 2, pp. 1–55, 2025
2025
Closest in time.
2025
Closest in time.