Fetching the paper…
Reading the bibliography…
Symmetry, a fundamental concept to understand our environment, often oversimplifies reality from a mathematical perspective.
F. M. Jaeger, Lectures on the principle of symmetry and its applications in all natural sciences, Elsevier, Amsterdam, 1917
1917
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, K. Kavukcuoglu, Asynchronous methods for deep reinforcement learning, in: International conference on machine learning, Vol. 48, 2016, pp. 1928–1937
1937
Earlier work this paper cites.
H. Weyl, Symmetry, Princeton University Press, Princeton, 1952
1952
Earlier work this paper cites.
R. McWeeny, Symmetry: an introduction to group theory and its applications, Pergamon Press, MacMillan, New York, 1963
1963
Earlier work this paper cites.
L.-J. Lin, Self-improving reactive agents based on reinforcement learning, planning and teaching, Machine learning 8 (3-4) (1992) 293–321
1992
Earlier work this paper cites.
C. J. C. H. Watkins, P. Dayan, Q-learning, Machine learning 8 (3) (1992) 279–292
1992
Earlier work this paper cites.
B. Ravindran, A. G. Barto, Symmetries and model minimization in Markov decision processes, Tech. rep., University of Massachusetts, USA (2001)
2001
Earlier work this paper cites.
M. Zinkevich, T. R. Balch, Symmetry in Markov decision processes and its implications for single agent and multiagent learning, in: Proceedings of the Eighteenth International Conference on Machine Learning, ICML ’01, Morgan Kaufmann Publishers Inc., San Francisco, CA, USA, 2001, p. 632
2001
Earlier work this paper cites.
K. Mainzer, Symmetry and complexity: The spirit and beauty of nonlinear science, Vol. 51, World Scientific, 2005
2005
Earlier work this paper cites.
A. Agostini, E. Celaya, Exploiting domain symmetries in reinforcement learning with continuous state and action spaces, in: 2009 International Conference on Machine Learning and Applications, IEEE, 2009, pp. 331–336
2009
Earlier work this paper cites.
B. S. Everitt, A. Skrondal, The Cambridge dictionary of statistics, 4th Edition, Cambridge University Press, Cambridge, UK, 2010
2010
Earlier work this paper cites.
I. Handžić, K. B. Reed, Perception of gait patterns that deviate from normal and symmetric biped locomotion, Frontiers in psychology 6 (2015)
2015
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, P. Moritz, Trust region policy optimization, in: Proceedings of the 32nd International Conference on Machine Learning, Vol. 37 of Proceedings of Machine Learning Research, PMLR, 2015, pp. 1889–1897
2015
Earlier work this paper cites.
D. P. Kingma, J. Ba, Adam: A method for stochastic optimization, in: Y. Bengio, Y. LeCun (Eds.), 3rd International Conference on Learning Representations (ICLR), San Diego, CA, USA, 2015
2015
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, P. Moritz, Trust region policy optimization, in: Proceedings of the 32nd International Conference on Machine Learning, Vol. 37 of Proceedings of Machine Learning Research, PMLR, 2015, pp. 1889–1897
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al., Mastering the game of Go with deep neural networks and tree search, Nature 529 (7587) (2016) 484–489
2016
Cited alongside, same era.
T. Cohen, M. Welling, Group equivariant convolutional networks, in: Proceedings of The 33rd International Conference on Machine Learning, Vol. 48 of Proceedings of Machine Learning Research, PMLR, 2016, pp. 2990–2999
2016
Cited alongside, same era.
2017
Cited alongside, same era.
X. B. Peng, G. Berseth, K. Yin, M. van de Panne, Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning, ACM Transactions on Graphics (Proc. SIGGRAPH 2017) 36 (4) (2017)
2017
Cited alongside, same era.
M. Papadatou-Pastou, E. Ntolka, J. Schmitz, M. Martin, M. R. Munafò, S. Ocklenburg, S. Paracchini, Human handedness: A meta-analysis, Psychological bulletin 146 (6) (2020) 481–524
2020
Later among the works it cites.
Z. Xie, P. Clary, J. Dao, P. Morais, J. Hurst, M. van de Panne, Learning locomotion skills for Cassie: Iterative design and sim-to-real, in: Proceedings of the Conference on Robot Learning, Vol. 100 of Proceedings of Machine Learning Research, PMLR, 2020, pp. 317–329
2020
Later among the works it cites.
Y. Lin, J. Huang, M. Zimmer, Y. Guan, J. Rojas, P. Weng, Invariant transform experience replay: Data augmentation for deep reinforcement learning, IEEE Robotics and Automation Letters 5 (4) (2020) 6615–6622
2020
Later among the works it cites.
E. van der Pol, D. E. Worrall, H. van Hoof, F. A. Oliehoek, M. Welling, MDP homomorphic networks: Group symmetries in reinforcement learning, in: Proceedings of the 34th International Conference on Neural Information Processing Systems, 2020, pp. 4199–4210
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Ravanbakhsh, J. Schneider, B. Póczos, Equivariance through parameter-sharing, in: Proceedings of the 34th International Conference on Machine Learning, Vol. 70 of Proceedings of Machine Learning Research, PMLR, 2017, pp. 2892–2901
2017
Cited alongside, same era.
A. Mahajan, T. Tulabandhula, Symmetry detection and exploitation for function approximation in deep RL, in: Proceedings of the 16th Conference on Autonomous Agents and MultiAgent Systems, AAMAS ’17, International Foundation for Autonomous Agents and Multiagent Systems, 2017, p. 1619–1621
2017
Cited alongside, same era.
2017
Cited alongside, same era.
P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, Y. Wu, P. Zhokhov, OpenAI baselines, https://github.com/openai/baselines (2017)
2017
Cited alongside, same era.
W. Yu, G. Turk, C. K. Liu, Learning symmetric and low-energy locomotion, ACM Transactions on Graphics 37 (4) (July 2018)
2018
Cited alongside, same era.
A. Hereid, C. M. Hubicki, E. A. Cousineau, A. D. Ames, Dynamic humanoid locomotion: A scalable formulation for HZD gait optimization, IEEE Transactions on Robotics 34 (2) (2018) 370–387
2018
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, S. Levine, Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, in: International Conference on Machine Learning, PMLR, 2018, pp. 1861–1870
2018
Cited alongside, same era.
D. Surovik, K. Wang, M. Vespignani, J. Bruce, K. E. Bekris, Adaptive tensegrity locomotion: Controlling a compliant icosahedron with symmetry-reduced reinforcement learning, The International Journal of Robotics Research (2019)
2019
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
A. Raffin, RL baselines3 zoo, https://github.com/DLR-RM/rl-baselines3-zoo (2020)
2020
Later among the works it cites.
M. G. Browne, C. S. Smock, R. T. Roemmich, The human preference for symmetric walking often disappears when one leg is constrained, The Journal of Physiology 599 (4) (2021) 1243–1260
2021
Later among the works it cites.
2021
Later among the works it cites.
K. Zeng, M. D. Graham, Symmetry reduction for deep reinforcement learning active control of chaotic spatiotemporal dynamics, Physical Review E 104 (July 2021)
2021
Later among the works it cites.
P. Ildefonso, P. Remédios, R. Silva, M. Vasco, F. S. Melo, A. Paiva, M. Veloso, Exploiting symmetry in human robot-assisted dressing using reinforcement learning, in: Progress in Artificial Intelligence: 20th EPIA Conference on Artificial Intelligence, Vol. 12981 of Lecture Notes in Computer Science, Springer, 2021, pp. 405–417
2021
Later among the works it cites.
R. van Bree, Data augmentation for regularizing learned world models in reinforcement learning, Master’s thesis, University of Twente (2021)
2021
Later among the works it cites.
A. Raffin, A. Hill, A. Gleave, A. Kanervisto, M. Ernestus, N. Dormann, Stable-Baselines3: Reliable reinforcement learning implementations (2021)
2021
Later among the works it cites.
M. Andrychowicz, A. Raichuk, P. Stańczyk, M. Orsini, S. Girgin, R. Marinier, L. Hussenot, …, O. Bachem, What matters for on-policy deep actor-critic methods? a large-scale study, in: International Conference on Learning Representations (ICLR), 2021
2021
Later among the works it cites.
A. Bhattacharya, M. Mattheakis, P. Protopapas, Encoding involutory invariances in neural networks, in: 2022 International Joint Conference on Neural Networks (IJCNN), 2022
2022
Later among the works it cites.
E. Coumans, Y. Bai, PyBullet, a Python module for physics simulation for games, robotics and machine learning, http://pybullet.org (2016–2024)
2024
Closest in time.