Fetching the paper…
Reading the bibliography…
Emergent effects can arise in multi-agent systems (MAS) where execution is decentralized and reliant on local information.
Science 177(4047), 393–396 (1972)
Anderson, P.W.: More is different · 1972
Earlier work this paper cites.
Handbooks in operations research and management science 2, 331–434 (1990)
Puterman, M.L.: Markov decision processes · 1990
Earlier work this paper cites.
In: Machine learning proceedings 1994, pp. 157–163. Elsevier (1994)
Littman, M.L.: Markov games as a framework for multi-agent reinforcement learning · 1994
Earlier work this paper cites.
OUP Oxford (1995)
Honderich, T.: The Oxford companion to philosophy · 1995
Earlier work this paper cites.
In: Seminal graphics: pioneering efforts that shaped the field, pp. 1–6. Association for Computing Machinery (1998)
Bresenham, J.E.: Algorithm for computer control of a digital plotter · 1998
Earlier work this paper cites.
In: Proceedings of the Sixteenth International Conference on Machine Learning (ICML). pp. 278–287 (1999)
Ng, A., Harada, D., Russell, S.: Policy invariance under reward transformations: Theory and application to reward shaping · 1999
Earlier work this paper cites.
Advances in Complex Systems 04(02n03) (2001)
Wolpert, D.H., Tumer, K.: Optimal payoff functions for members of collectives · 2001
Earlier work this paper cites.
In: International conference on complex systems. vol. 21, pp. 16–21. Citeseer (2004)
Tisue, S., Wilensky, U.: Netlogo: A simple environment for modeling complexity · 2004
Earlier work this paper cites.
arXiv preprint nlin/0506028 (2005)
Fromm, J.: Types and forms of emergence · 2005
Earlier work this paper cites.
In: International Symposium on Formal Methods for Components and Objects. pp. 1–24 (2011)
Wirsing, M., Hölzl, M., Tribastone, M., Zambonelli, F.: ASCENS: Engineering autonomic service-component ensembles · 2011
Earlier work this paper cites.
Journal of Artificial Intelligence Research 48, 67–113 (2013)
Roijers, D.M., Vamplew, P., Whiteson, S., Dazeley, R.: A survey of multi-objective sequential decision-making · 2013
Earlier work this paper cites.
The Journal of Machine Learning Research 15(1), 3483–3512 (2014)
Van Moffaert, K., Nowé, A.: Multi-objective reinforcement learning using sets of pareto dominating policies · 2014
Earlier work this paper cites.
MIT press (2015)
Wilensky, U., Rand, W.: An introduction to agent-based modeling: modeling natural, social, and engineered complex systems with NetLogo · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1606.06565 (2016)
Amodei, D., Olah, C., Steinhardt, J., Christiano, P., Schulman, J., Mané, D.: Concrete problems in AI safety · 2016
Earlier work this paper cites.
In: International conference on machine learning. pp. 1928–1937. PMLR (2016)
Mnih, V., Badia, A.P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., Kavukcuoglu, K.: Asynchronous methods for deep reinforcement learning · 2016
Cited alongside, same era.
In: Advances in Neural Information Processing Systems (NeurIPS). vol. 30, p. 4302–4310 (2017)
Christiano, P.F., Leike, J., Brown, T., Martic, M., Legg, S., Amodei, D.: Deep reinforcement learning from human preferences · 2017
Cited alongside, same era.
preprint arxiv:1711.0988 (2017)
Leike, J., Martic, M., Krakovna, V., Ortega, P.A., Everitt, T., Lefrancq, A., Orseau, L., Legg, S.: AI safety gridworlds · 2017
Cited alongside, same era.
In: Proceedings of the Thirty-Second AAAI Conference on Artificial Intelligence. p. 2974–2982 (02 2018)
Foerster, J.N., Farquhar, G., Afouras, T., Nardelli, N., Whiteson, S.: Counterfactual multi-agent policy gradients · 2018
Cited alongside, same era.
Autonomous Agents and Multi-Agent Systems 36(1), 26 (2022)
Hayes, C.F., Rădulescu, R., Bargiacchi, E., Källström, J., Macfarlane, M., Reymond, M., Verstraeten, T., Zintgraf, L.M., Dazeley, R., Heintz, F., et al.: A practical guide to multi-objective reinforcement learning and planning · 2022
Later among the works it cites.
Standard, International Organization for Standardization (2022)
ISO21448:2022(E): Road vehicles — safety of the intended functionality · 2022
Later among the works it cites.
In: Advances in Neural Information Processing Systems (NeurIPS). vol. 35, pp. 27730–27744 (2022)
Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P.F., Leike, J., Lowe, R.: Training language models to follow instructions with human feedback · 2022
Later among the works it cites.
arXiv preprint arXiv:2204.05036 (2022)
Reymond, M., Bargiacchi, E., Nowé, A.: Pareto conditioned networks · 2022
Later among the works it cites.
In: Agents and Artificial Intelligence. pp. 3–21. Springer International Publishing (2022)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Leike, J., Krueger, D., Everitt, T., Martic, M., Maini, V., Legg, S.: Scalable agent alignment via reward modeling: a research direction · 2018
Cited alongside, same era.
arXiv preprint arXiv:1802.09464 (2018)
Plappert, M., Andrychowicz, M., Ray, A., McGrew, B., Baker, B., Powell, G., Schneider, J., Tobin, J., Chociej, M., Welinder, P., et al.: Multi-goal reinforcement learning: Challenging robotics environments and request for research · 2018
Cited alongside, same era.
MIT press (2018)
Sutton, R.S., Barto, A.G.: Reinforcement learning: An introduction · 2018
Cited alongside, same era.
Allen Lane (2019)
Russell, S.: Human Compatible: Artificial Intelligence and the Problem of Control · 2019
Cited alongside, same era.
arXiv preprint arXiv:1901.02219 (2019)
Sedlmeier, A., Gabor, T., Phan, T., Belzner, L., Linnhoff-Popien, C.: Uncertainty-based out-of-distribution detection in deep reinforcement learning · 2019
Cited alongside, same era.
Minds and Machines 30(3), 411–437 (09 2020)
Gabriel, I.: Artificial intelligence, values, and alignment · 2020
Cited alongside, same era.
DeepMind Blog 3 (2020)
Krakovna, V., Uesato, J., Mikulik, V., Rahtz, M., Everitt, T., Kumar, R., Kenton, Z., Leike, J., Legg, S.: Specification gaming: the flip side of AI ingenuity · 2020
Cited alongside, same era.
Artificial life 26(2), 274–306 (2020)
Lehman, J., Clune, J., Misevic, D., Adami, C., Altenberg, L., Beaulieu, J., Bentley, P.J., Bernard, S., Beslon, G., Bryson, D.M., et al.: The surprising creativity of digital evolution: A collection of anecdotes from the evolutionary computation and artificial life research communities · 2020
Cited alongside, same era.
Ritz, F., Phan, T., Müller, R., Gabor, T., Sedlmeier, A., Zeller, M., Wieghardt, J., Schmid, R., Sauer, H., Klein, C., Linnhoff-Popien, C.: Specification aware multi-agent reinforcement learning · 2022
Later among the works it cites.
Sensors 22(4) (2022)
Salimibeni, M., Mohammadi, A., Malekzadeh, P., Plataniotis, K.N.: Multi-agent reinforcement learning via adaptive Kalman temporal difference and successor representation · 2022
Later among the works it cites.
In: Elkind, E. (ed.) Proceedings of the Thirty-Second International Joint Conference on Artificial Intelligence, IJCAI-23. pp. 3414–3422 (2023)
Altmann, P., Ritz, F., Feuchtinger, L., Nüßlein, J., Linnhoff-Popien, C., Phan, T.: CROP: Towards distributional-shift robust reinforcement learning using compact reshaped observation processing · 2023
Later among the works it cites.
Frontiers in Computer Science 5, 1132580 (2023)
Burton, S., Herd, B.: Addressing uncertainty in the safety assurance of machine-learning · 2023
Later among the works it cites.
In: Proceedings of the 37th Conference on Neural Information Processing Systems (NeurIPS 2023) (2023)
Felten, F., Alegre, L.N., Nowé, A., Bazzan, A.L.C., Talbi, E.G., Danoy, G., Silva, B.C.d.: A toolkit for reliable benchmarking and research in multi-objective reinforcement learning · 2023
Later among the works it cites.
In: Proceedings of the 2023 International Conference on Autonomous Agents and Multiagent Systems. pp. 851–859 (2023)
Haider, T., Roscher, K., Schmoeller da Roza, F., Günnemann, S.: Out-of-distribution detection for reinforcement learning agents with probabilistic dynamics models · 2023
Later among the works it cites.
In: Proceedings of the 2023 International Conference on Autonomous Agents and Multiagent Systems. pp. 1988–1990 (2023)
Hayes, C.F., Rădulescu, R., Bargiacchi, E., Kallstrom, J., Macfarlane, M., Reymond, M., Verstraeten, T., Zintgraf, L.M., Dazeley, R., Heintz, F., et al.: A brief guide to multi-objective reinforcement learning and planning · 2023
Later among the works it cites.
Towers, M., Terry, J.K., Kwiatkowski, A., Balis, J.U., de Cola, G., Deleu, T., Goulão, M., Kallinteris, A., KG, A., Krimmel, M., Perez-Vicente, R., Pierré, A., Schulhoff, S., Tai, J.J., Tan, A.J.S., Younis, O.G.: Gymnasium (2023), https://github.com/Farama-Foundation/Gymnasium
2023
Later among the works it cites.
In: Proceedings of the 39th ACM/SIGAPP Symposium on Applied Computing. pp. 1569–1578 (2024)
Haider, T., Roscher, K., Herd, B., Schmoeller Roza, F., Burton, S.: Can you trust your agent? The effect of out-of-distribution detection on the safety of reinforcement learning systems · 2024
Closest in time.