Fetching the paper…
Reading the bibliography…
This paper introduces Multi-Agent MDP Homomorphic Networks, a class of networks that allows distributed execution using only local information, yet is able to share experience between global symmetries in the joint state-action space of cooperative multi-agent systems.
Planning, learning and coordination in multiagent decision processes
Craig Boutilier · 1996
Earlier work this paper cites.
Symmetries and model minimization in Markov decision processes
Balaraman Ravindran and Andrew G. Barto · 2001
Earlier work this paper cites.
Collaborative multiagent reinforcement learning by payoff propagation
Jelle R Kok and Nikos Vlassis · 2006
Earlier work this paper cites.
Multiagent reinforcement learning for urban traffic control using coordination graphs
Lior Kuyer, Shimon Whiteson, Bram Bakker, and Nikos Vlassis · 2008
Earlier work this paper cites.
Optimal and approximate Q-value functions for decentralized POMDPs
Frans A Oliehoek, Matthijs TJ Spaan, and Nikos Vlassis · 2008
Earlier work this paper cites.
Symmetry in Markov decision processes and its implications for single agent and multi agent learning
Martin Zinkevich and Tucker Balch · 2008
Earlier work this paper cites.
Decentralized stochastic planning with anonymity in interactions
Pradeep Varakantham, Yossiri Adulyasak, and Patrick Jaillet · 2014
Earlier work this paper cites.
Teaching deep convolutional neural networks to play Go
Christopher Clark and Amos Storkey · 2015
Earlier work this paper cites.
Group equivariant convolutional networks
Taco Cohen and Max Welling · 2016
Earlier work this paper cites.
Multi-agent reinforcement learning as a rehearsal for decentralized planning
Landon Kraemer and Bikramjit Banerjee · 2016
Earlier work this paper cites.
Exploiting anonymity in approximate linear programming: Scaling to large multiagent MDPs
Philipp Robbel, Frans A Oliehoek, Mykel J Kochenderfer, et al · 2016
Earlier work this paper cites.
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar, Arthur Szlam, and Rob Fergus · 2016
Earlier work this paper cites.
Coordinated deep reinforcement learners for traffic light control
Elise van der Pol and Frans A Oliehoek · 2016
Earlier work this paper cites.
Steerable CNNs
Taco S. Cohen and Max Welling · 2017
Earlier work this paper cites.
Automatic differentiation in PyTorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Value-decomposition networks for cooperative multi-agent learning
Peter Sunehag, Guy Lever, Audrunas Gruslys, Wojciech Marian Czarnecki, Vinicius Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z Leibo, Karl Tuyls, et al · 2017
Cited alongside, same era.
Harmonic networks: Deep translation and rotation equivariance
Daniel E. Worrall, Stephan J. Garbin, Daniyar Turmukhambetov, and Gabriel J. Brostow · 2017
Cited alongside, same era.
Tensor field networks: Rotation-and translation-equivariant neural networks for 3d point clouds
Deep coordination graphs
Wendelin Böhmer, Vitaly Kurin, and Shimon Whiteson · 2020
Later among the works it cites.
SE (3)-transformers: 3d roto-translation equivariant attention networks
Fabian B Fuchs, Daniel E Worrall, Volker Fischer, and Max Welling · 2020
Later among the works it cites.
Array programming with NumPy
Charles R. Harris, K. Jarrod Millman, Stéfan J van der Walt, Ralf Gommers, Pauli Virtanen, David Cournapeau, Eric Wieser, Julian Taylor, Sebastian Berg, Nathaniel J. Smith, Robert Kern, Matti Picus, Stephan Hoyer, Marten H. van Kerkwijk, Matthew Brett, Allan Haldane, Jaime Fernández del Río, Mark Wiebe, Pearu Peterson, Pierre Gérard-Marchant, Kevin Sheppard, Tyler Reddy, Warren Weckesser, Hameer Abbasi, Christoph Gohlke, and Travis E. Oliphant · 2020
Later among the works it cites.
"Other-play" for zero-shot coordination
Hengyuan Hu, Adam Lerer, Alex Peysakhovich, and Jakob Foerster · 2020
Later among the works it cites.
Graph convolutional reinforcement learning
Jiechuan Jiang, Chen Dun, Tiejun Huang, and Zongqing Lu · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nathaniel Thomas, Tess Smidt, Steven Kearnes, Lusann Yang, Li Li, Kai Kohlhoff, and Patrick Riley · 2018
Cited alongside, same era.
3D steerable CNNs: Learning rotationally equivariant features in volumetric data
Maurice Weiler, Mario Geiger, Max Welling, Wouter Boomsma, and Taco Cohen · 2018
Cited alongside, same era.
3D G-CNNs for pulmonary nodule detection
Marysia Winkels and Taco S. Cohen · 2018
Cited alongside, same era.
On learning symmetric locomotion
Farzad Abdolhosseini, Hung Yu Ling, Zhaoming Xie, Xue Bin Peng, and Michiel van de Panne · 2019
Cited alongside, same era.
PIC: Permutation invariant critic for multi-agent deep reinforcement learning
Iou-Jen Liu, Raymond A. Yeh, and Alexander G. Schwing · 2019
Cited alongside, same era.
Augmenting learning using symmetry in a biologically-inspired domain
Shruti Mishra, Abbas Abdolmaleki, Arthur Guez, Piotr Trochim, and Doina Precup · 2019
Cited alongside, same era.
rlpyt: A research code base for deep reinforcement learning in Pytorch
Adam Stooke and Pieter Abbeel · 2019
Cited alongside, same era.
A survey on traffic signal control methods
Hua Wei, Guanjie Zheng, Vikash Gayah, and Zhenhui Li · 2019
Cited alongside, same era.
Michael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto, Pieter Abbeel, and Aravind Srinivas · 2020
Later among the works it cites.
Invariant transform experience replay: Data augmentation for deep reinforcement learning
Yijiong Lin, Jiancong Huang, Matthieu Zimmer, Yisheng Guan, Juan Rojas, and Paul Weng · 2020
Later among the works it cites.
Goal-conditioned batch reinforcement learning for rotation invariant locomotion
Aditi Mavalankar · 2020
Later among the works it cites.
MDP homomorphic networks: Group symmetries in reinforcement learning
Elise van der Pol, Daniel E. Worrall, Herke van Hoof, Frans Oliehoek, and Max Welling · 2020
Later among the works it cites.
Group equivariant generative adversarial networks
Neel Dey, Antong Chen, and Soheil Ghafurian · 2021
Closest in time.
Image augmentation is all you need: Regularizing deep reinforcement learning from pixels
Ilya Kostrikov, Denis Yarats, and Rob Fergus · 2021
Closest in time.
Neural enhanced belief propagation on factor graphs
Victor Garcia Satorras and Max Welling · 2021
Closest in time.
Symmetry-aware actor-critic for 3d molecular design
Gregor NC Simm, Robert Pinsler, Gábor Csányi, and José Miguel Hernández-Lobato · 2021
Closest in time.