Fetching the paper…
Reading the bibliography…
We propose a model-free algorithm for learning efficient policies capable of returning table tennis balls by controlling robot joints at a rate of 100Hz.
Robot ping pong
John Billingsley · 1983
Earlier work this paper cites.
Toshiba progress towards sensory control in real time
J. Hartley · 1983
Earlier work this paper cites.
Pingpong-playing robot controlled by a microcomputer
John Knight and David Lowery · 1986
Earlier work this paper cites.
Development of ping-pong robot system using 7 degree of freedom direct drive robots
Hideaki Hashimoto, Fumio Ozaki, and Kuniji Osuka · 1987
Earlier work this paper cites.
A Robot Ping-Pong Player: Experiments in Real-Time Intelligent Control
Russell Anderson · 1988
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald Williams · 1992
Earlier work this paper cites.
Towards a stroke construction model
Marie-Marin Ramanantsoa and Alain Durey · 1994
Earlier work this paper cites.
Reinforcement Learning: An introduction
Richard Sutton and Andrew Barto · 1998
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard Sutton, David McAllester, Satinder Singh, and Yishay Mansour · 2000
Earlier work this paper cites.
Movement imitation with nonlinear dynamical systems in humanoid robots
Auke Jan Ijspeert, Jun Nakanishi, and Stefan Schaal · 2002
Earlier work this paper cites.
Realization of the table tennis task based on virtual targets
Fumio Miyazaki, Masahiro Takeuchi, Michiya Matsushima, Takamichi Kusano, and Takaaki Hashimoto · 2002
Earlier work this paper cites.
Learning to the robot table tennis task-ball control and rally with a human
Michiya Matsushima, Takaaki Hashimoto, and Fumio Miyazaki · 2003
Earlier work this paper cites.
A learning approach to robotic table tennis
Michiya Matsushima, Takaaki Hashimoto, Masahiro Takeuchi, and Fumio Miyazaki · 2005
Earlier work this paper cites.
Learning to dynamically manipulate: A table tennis robot controls a ball and rallies with a human being
Fumio Miyazaki et al · 2006
Earlier work this paper cites.
Natural evolution strategies
Daan Wierstra, Tom Schaul, Jan Peters, and Jürgen Schmidhuber · 2008
Cited alongside, same era.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston · 2009
Cited alongside, same era.
A computational model of human table tennis for robot application
Katharina Muelling and Jan Peters · 2009
Cited alongside, same era.
Balance motion generation for a humanoid robot playing table tennis
Yichao Sun, Rong Xiong, Qiuguo Zhu, Jingjing Wu, and Jian Chu · 2011
Cited alongside, same era.
Learning to select and generalize striking movements in robot table tennis
Katharina Muelling, Jens Kober, Oliver Kroemer, and Jan Peters · 2012
Cited alongside, same era.
Learning optimal striking points for a ping-pong playing robot
Yanlong Huang, Bernhard Schölkopf, and Jan Peters · 2015
Cited alongside, same era.
Structured evolution with compact architectures for scalable policy optimization
Krzysztof Choromanski et al · 2018
Later among the works it cites.
Meta-learning and universality: Deep representations and gradient descent can approximate any learning algorithm
Chelsea Finn and Sergey Levine · 2018
Later among the works it cites.
Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation
Dmitry Kalashnikov et al · 2018
Later among the works it cites.
Online optimal trajectory generation for robot table tennis
Okan Koç, Guilherme Maeda, and Jan Peters · 2018
Later among the works it cites.
Hierarchical policy design for sample-efficient learning of robot table tennis through self-play
Reza Mahjourian, Navdeep Jaitly, Nevena Lazic, Sergey Levine, and Risto Miikkulainen · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Model-free trajectory-based policy optimization with monotonic improvement
Riad Akrour, Abbas Abdolmaleki, Hany Abdulsamad, Jan Peters, and Gerhard Neumann · 2016
Cited alongside, same era.
Jointly learning trajectory generation and hitting point prediction in robot table tennis
Yanlong Huang, Dieter Buchler, Okan Koç, Bernhard Schölkopf, and Jan Peters · 2016
Cited alongside, same era.
Continuous control with deep reinforcement learning
Timothy Lillicrap, Jonathan Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2016
Cited alongside, same era.
Conditional image generation with pixelcnn decoders
Aaron van den Oord, Nal Kalchbrenner, Lasse Espeholt, koray kavukcuoglu, Oriol Vinyals, and Alex Graves · 2016
Cited alongside, same era.
Random gradient-free minimization of convex functions
Yurii Nesterov and Vladimir Spokoiny · 2017
Cited alongside, same era.
Towards generalization and simplicity in continuous control
Aravind Rajeswaran, Kendall Lowrey, Emanuel Todorov, and Sham M. Kakade · 2017
Cited alongside, same era.
Simple random search provides a competitive approach to reinforcement learning
Horia Mania, Aurelia Guy, and Benjamin Recht · 2018
Later among the works it cites.
Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning
Anusha Nagabandi et al · 2018
Later among the works it cites.
A table tennis robot system using an industrial kuka robot arm
Jonas Tebbe, Yapeng Gao, Marc Sastre-Rienietz, and Andreas Zell · 2018
Later among the works it cites.
Domain randomization and generative models for robotic grasping
Josh Tobin et al · 2018
Later among the works it cites.
Towards high level skill learning: Learn to return table tennis ball using monte-carlo based policy gradient method
Yifeng Zhu, Yongsheng Zhao, Lisen Jin, Jingjing Wu, and Rong Xiong · 2018
Later among the works it cites.
Pybullet, a python module for physics simulation for games, robotics and machine learning
Erwin Coumans and Yunfei Bai · 2019
Later among the works it cites.
Markerless racket pose detection and stroke classification based on stereo vision for table tennis robots
Yapeng Gao, Jonas Tebbe, Julian Krismer, and Andreas Zell · 2019
Later among the works it cites.
Sim-to-real via sim-to-sim: Data-efficient robotic grasping via randomized-to-canonical adaptation networks
Stephen James, Paul Wohlhart, Mrinal Kalakrishnan, Dmitry Kalashnikov, Alex Irpan, Julian Ibarz, Sergey Levine, Raia Hadsell, and Konstantinos Bousmalis · 2019
Later among the works it cites.
ES-MAML: simple hessian-free meta learning
Xingyou Song, Wenbo Gao, Yuxiang Yang, Krzysztof Choromanski, Aldo Pacchiano, and Yunhao Tang · 2020
Closest in time.