Fetching the paper…
Reading the bibliography…
Modeling agent behavior is central to understanding the emergence of complex phenomena in multiagent systems.
Efficient training of artificial neural networks for autonomous navigation
Pomerleau, D. A · 1991
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Littman, M. L · 1994
Earlier work this paper cites.
Multi-agent systems: An introduction to distributed artificial intelligence , volume 1
Ferber, J · 1999
Earlier work this paper cites.
Active learner modelling
McCalla, G., Vassileva, J., Greer, J., and Bull, S · 2000
Earlier work this paper cites.
Multiagent systems: A survey from a machine learning perspective
Stone, P. and Veloso, M · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P. and Ng, A. Y · 2004
Earlier work this paper cites.
Learning deep architectures for ai
Bengio, Y. et al · 2009
Earlier work this paper cites.
An introduction to multiagent systems
Wooldridge, M · 2009
Earlier work this paper cites.
Computer poker: A review
Rubin, J. and Watson, I · 2011
Earlier work this paper cites.
MuJoCo: A physics engine for model-based control
Todorov, E., Erez, T., and Tassa, Y · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Mikolov, T., Sutskever, I., Chen, K., Corrado, G. S., and Dean, J · 2013
Cited alongside, same era.
Deep metric learning using triplet network
Hoffer, E. and Ailon, N · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al · 2015
Cited alongside, same era.
Interaction networks for learning about objects, relations and physics
Battaglia, P., Pascanu, R., Lai, M., Rezende, D. J., et al · 2016
Cited alongside, same era.
node2vec: Scalable feature learning for networks
Grover, A. and Leskovec, J · 2016
Cited alongside, same era.
Coordinated multi-agent imitation learning
Le, H. M., Yue, Y., and Carr, P · 2017
Later among the works it cites.
Inferring the latent structure of human decision-making from raw visual inputs
Li, Y., Song, J., and Ermon, S · 2017
Later among the works it cites.
Multi-agent actor-critic for mixed cooperative-competitive environments
Lowe, R., Wu, Y., Tamar, A., Harb, J., Abbeel, P., and Mordatch, I · 2017
Later among the works it cites.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Later among the works it cites.
Robust imitation of diverse behaviors
Wang, Z., Merel, J., Reed, S., Wayne, G., de Freitas, N., and Heess, N · 2017
Later among the works it cites.
Continuous adaptation via meta-learning in nonstationary and competitive environments
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Autonomous agents modeling other agents: A comprehensive survey and open problems
Albrecht, S. V. and Stone, P · 2017
Cited alongside, same era.
Multiplayer reach-avoid games via pairwise outcomes
Chen, M., Zhou, Z., and Tomlin, C. J · 2017
Cited alongside, same era.
VAIN: Attentional multi-agent predictive modeling
Hoshen, Y · 2017
Cited alongside, same era.
Dynamics on linear influence network games under stochastic environments
Zhou, Z., Bambos, N., and Glynn, P
Cited in the paper.
A game-theoretical formulation of influence networks
Zhou, Z., Yolken, B., Miura-Ko, R. A., and Bambos, N
Cited in the paper.
Al-Shedivat, M., Bansal, T., Burda, Y., Sutskever, I., Mordatch, I., and Abbeel, P · 2018
Closest in time.
Evaluating generalization in multiagent systems using agent-interaction graphs
Grover, A., Al-Shedivat, M., Gupta, J. K., Burda, Y., and Edwards, H · 2018
Closest in time.
Neural relational inference for interacting systems
Kipf, T., Fetaya, E., Wang, K.-C., Welling, M., and Zemel, R · 2018
Closest in time.
Emergence of grounded compositional language in multi-agent populations
Mordatch, I. and Abbeel, P · 2018
Closest in time.