Fetching the paper…
Reading the bibliography…
Multi-agent systems represent a significant advancement in artificial intelligence, enabling complex problem-solving through coordinated specialized agents.
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., … & Amodei, D. (2023). Language models are few-shot learners. Advances in Neural Information Processing Systems
1901
Earlier work this paper cites.
1907
Earlier work this paper cites.
Kuhn, H. W. (1955). The Hungarian method for the assignment problem. Naval Research Logistics Quarterly
1955
Earlier work this paper cites.
Kahneman, D., & Tversky, A. (1979). Prospect theory: An analysis of decision under risk. Econometrica
1979
Earlier work this paper cites.
Langley, P. (1995). Order effects in incremental learning. In Learning in Humans and Machines: Towards an Interdisciplinary Learning Science
1995
Earlier work this paper cites.
2001
Earlier work this paper cites.
Foerster, J., Farquhar, G., Afouras, T., Nardelli, N., & Whiteson, S. (2024). Counterfactual multi-agent policy gradients. In Proceedings of the 41st International Conference on Machine Learning
2004
Earlier work this paper cites.
Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., … & Zheng, X. (2023). TensorFlow: Large-scale machine learning on heterogeneous systems. Journal of Machine Learning Research
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Andrychowicz, M., Baker, B., Chociej, M., Jozefowicz, R., McGrew, B., Pachocki, J., … & Zaremba, W. (2023). Learning dexterous in-hand manipulation. The International Journal of Robotics Research
2023
Earlier work this paper cites.
Arrieta, A. B., Díaz-Rodríguez, N., Del Ser, J., Bennetot, A., Tabik, S., Barbado, A., … & Herrera, F. (2023). Explainable Artificial Intelligence (XAI): Concepts, taxonomies, opportunities and challenges toward responsible AI. Information Fusion
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Christiano, P. F., Leike, J., Brown, T., Martic, M., Legg, S., & Amodei, D. (2023). Deep reinforcement learning from human preferences. Advances in Neural Information Processing Systems
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Devlin, J., Chang, M. W., Lee, K., & Toutanova, K. (2023). BERT: Pre-training of deep bidirectional transformers for language understanding. Computational Linguistics
2023
Earlier work this paper cites.
Doshi-Velez, F., & Kim, B. (2023). Considerations for evaluation of explainable artificial intelligence methods. Nature Machine Intelligence
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Dulac-Arnold, G., Mankowitz, D., & Hester, T. (2023). Challenges of real-world reinforcement learning. Journal of Artificial Intelligence Research
2023
Earlier work this paper cites.
Dwivedi, Y. K., Hughes, L., Ismagilova, E., Aarts, G., Coombs, C., Crick, T., … & Williams, M. D. (2023). Artificial Intelligence (AI): Multidisciplinary perspectives on emerging challenges, opportunities, and agenda for research, practice and policy. International Journal of Information Management
2023
Earlier work this paper cites.
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., … & Kavukcuoglu, K. (2023). IMPALA: Scalable distributed deep-RL with importance weighted actor-learner architectures. In Proceedings of the 40th International Conference on Machine Learning
2023
Earlier work this paper cites.
Fawzi, A., Balog, M., Huang, A., Hubert, T., Romera-Paredes, B., Barekatain, M., … & Kohli, P. (2023). Discovering faster matrix multiplication algorithms with reinforcement learning. Nature
2023
Earlier work this paper cites.
Finn, C., Abbeel, P., & Levine, S. (2023). Model-agnostic meta-learning for fast adaptation of deep networks. Journal of Machine Learning Research
2023
Earlier work this paper cites.
Geirhos, R., Jacobsen, J. H., Michaelis, C., Zemel, R., Brendel, W., Bethge, M., & Wichmann, F. A. (2023). Shortcut learning in deep neural networks. Nature Machine Intelligence
2023
Earlier work this paper cites.
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., … & Bengio, Y. (2023). Generative adversarial networks. Communications of the ACM
2023
Earlier work this paper cites.
Graves, A., Wayne, G., Reynolds, M., Harley, T., Danihelka, I., Grabska-Barwińska, A., … & Hassabis, D. (2023). Hybrid computing using a neural network with dynamic external memory. Nature
2023
Earlier work this paper cites.
Ha, D., & Schmidhuber, J. (2023). World models. Nature Machine Intelligence
2023
Earlier work this paper cites.
Hadfield-Menell, D., Dragan, A., Abbeel, P., & Russell, S. (2023). Cooperative inverse reinforcement learning. Advances in Neural Information Processing Systems
2023
Cited alongside, same era.
Hendrycks, D., Burns, C., Basart, S., Zou, A., Mazeika, M., Song, D., & Steinhardt, J. (2023). Measuring massive multitask language understanding. In Proceedings of the International Conference on Learning Representations (ICLR)
2023
Cited alongside, same era.
Hoffmann, J., Borgeaud, S., Mensch, A., Buchatskaya, E., Cai, T., Rutherford, E., … & Sifre, L. (2023). Training compute-optimal large language models. Advances in Neural Information Processing Systems
2023
Cited alongside, same era.
Jaderberg, M., Czarnecki, W. M., Dunning, I., Marris, L., Lever, G., Castaneda, A. G., … & Graepel, T. (2023). Human-level performance in 3D multiplayer games with population-based reinforcement learning. Science
2023
Cited alongside, same era.
Chen, L., Lu, K., Rajeswaran, A., Lee, K., Grover, A., Laskin, M., … & Mordatch, I. (2024). Decision transformer: Reinforcement learning via sequence modeling. Journal of Machine Learning Research
2024
Later among the works it cites.
Choi, E., Hewitt, J., Uszkoreit, J., Lacoste, A., Larochelle, H., Charton, F., & Mathis, M. (2024). On the opportunities and challenges of foundation models for geoscience. Nature Machine Intelligence
2024
Later among the works it cites.
Cranmer, M. D., Xu, R., Battaglia, P., & Ho, S. (2024). Learning symbolic physics with graph networks. Nature Machine Intelligence
2024
Later among the works it cites.
Deng, S., Zhao, H., Yin, J., Dustdar, S., & Zomaya, A. Y. (2024). Edge intelligence: The confluence of edge computing and artificial intelligence. IEEE Internet of Things Journal
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jang, E., Gu, S., & Poole, B. (2023). Categorical reparameterization with Gumbel-Softmax. In Proceedings of the International Conference on Learning Representations (ICLR)
2023
Cited alongside, same era.
Jumper, J., Evans, R., Pritzel, A., Green, T., Figurnov, M., Ronneberger, O., … & Hassabis, D. (2023). Highly accurate protein structure prediction with AlphaFold. Nature
2023
Cited alongside, same era.
Kaelbling, L. P., Littman, M. L., & Moore, A. W. (2023). Reinforcement learning: A survey. Journal of Artificial Intelligence Research
2023
Cited alongside, same era.
Karpathy, A., & Fei-Fei, L. (2023). Deep visual-semantic alignments for generating image descriptions. IEEE Transactions on Pattern Analysis and Machine Intelligence
2023
Cited alongside, same era.
Kingma, D. P., & Welling, M. (2023). Auto-encoding variational Bayes. Journal of Machine Learning Research
2023
Cited alongside, same era.
Kober, J., Bagnell, J. A., & Peters, J. (2023). Reinforcement learning in robotics: A survey. The International Journal of Robotics Research
2023
Cited alongside, same era.
Kottur, S., Moura, J. M., Lee, S., & Batra, D. (2023). Natural language does not emerge ’naturally’ in multi-agent dialog. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing
2023
Cited alongside, same era.
Krizhevsky, A., Sutskever, I., & Hinton, G. E. (2023). ImageNet classification with deep convolutional neural networks. Communications of the ACM
2023
Cited alongside, same era.
Du, N., Huang, Y., Dai, A. M., Tong, S., Lepikhin, D., Xu, Y., … & Dean, J. (2024). GLaM: Efficient scaling of language models with mixture-of-experts. Journal of Machine Learning Research
2024
Later among the works it cites.
Dubey, R., Agrawal, P., Pathak, D., Griffiths, T. L., & Efros, A. A. (2024). Investigating human priors for playing video games. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.
Durugkar, I., Rosenbaum, C., Dernbach, S., & Raghavan, V. (2024). Deep multi-agent reinforcement learning for decentralized continuous cooperative control. IEEE Transactions on Neural Networks and Learning Systems
2024
Later among the works it cites.
2024
Later among the works it cites.
Eysenbach, B., Salakhutdinov, R., & Levine, S. (2024). Search on the replay buffer: Bridging planning and reinforcement learning. Advances in Neural Information Processing Systems
2024
Later among the works it cites.
2024
Later among the works it cites.
Gerstgrasser, M., Vogel, D., & Heess, N. (2024). Deep reinforcement learning with relational inductive biases. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.
Guo, M., Zhang, Y., & Liu, T. (2024). Efficient training of language models to fill in the middle. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.
Hao, J., Leung, H. F., & Ming, Z. (2024). Autonomous agents: Past, present, and future. IEEE Transactions on Knowledge and Data Engineering
2024
Later among the works it cites.
Hernandez, D., Kaplan, J., Henighan, T., & McCandlish, S. (2024). Scaling laws for transfer. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.
Huang, S., Ott, M., Auli, M., & Liu, H. (2024). Improving language model behavior by training on a curated dataset. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.
Jain, A., Liu, I., Lazaridou, A., & Graepel, T. (2024). Improving coordination in multi-agent reinforcement learning through self-supervision. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Koh, P. W., Sagawa, S., Marklund, H., Xie, S. M., Zhang, M., Balsubramani, A., … & Liang, P. (2024). WILDS: A benchmark of in-the-wild distribution shifts. Journal of Machine Learning Research
2024
Later among the works it cites.
Kumar, A., Fu, J., Tucker, G., & Levine, S. (2024). Stabilizing off-policy Q-learning via bootstrapping error reduction. Advances in Neural Information Processing Systems
2024
Later among the works it cites.
Lample, G., Sablayrolles, A., Rouditchenko, A., Noroozi, L., Houlsby, N., & Usunier, N. (2024). Hyena hierarchy: Towards larger convolutional language models. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.
Lee, J., Mansimov, E., & Cho, K. (2024). MIND2WEB: Towards a general-purpose agent for web navigation. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.
Li, Y., Choi, D., Chung, J., Kushman, N., Schrittwieser, J., Leblond, R., … & Vinyals, O. (2024). Competition-level code generation with AlphaCode. Science
2024
Later among the works it cites.
Lin, K., Li, D., He, X., Zhang, Z., & Sun, M. T. (2024). Adversarial multi-agent reinforcement learning with graph convolutional networks. IEEE Transactions on Neural Networks and Learning Systems
2024
Later among the works it cites.
Liu, H., Tam, D., Muqeeth, M., Mohta, J., Huang, T., Bansal, M., & Raffel, C. (2024). Lost in the middle: How language models use long contexts. In Proceedings of the 41st International Conference on Machine Learning
2024
Later among the works it cites.