Fetching the paper…
Reading the bibliography…
Action representation is an important yet often overlooked aspect in end-to-end robot learning with deep networks.
Y. LeCun, S. Chopra, R. Hadsell, M. Ranzato, and F. Huang, “A tutorial on energy-based learning,”
2006
Earlier work this paper cites.
D. Berenson, S. Srinivasa, and J. Kuffner, “Task space regions: A framework for pose-constrained manipulation planning,”
2011
Earlier work this paper cites.
M. Welling and Y. W. Teh, “Bayesian learning via stochastic gradient langevin dynamics,” in
2011
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep Residual Learning for Image Recognition,” in
2016
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end Training of Deep Visuomotor Policies,” in
2016
Earlier work this paper cites.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous Control with Deep Reinforcement Learning,” in
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
R. Fox, S. Krishnan, I. Stoica, and K. Goldberg, “DDCO: Discovery of Deep Continuous Options for Robot Learning from Demonstrations,” in
2017
Earlier work this paper cites.
X. B. Peng and M. van de Panne, “Learning locomotion skills using deeprl: Does the choice of action space matter?” in
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
T. Haarnoja, K. Hartikainen, P. Abbeel, and S. Levine, “Latent space policies for hierarchical reinforcement learning,” in
2018
Earlier work this paper cites.
K. Hausman, J. T. Springenberg, Z. Wang, N. Heess, and M. A. Riedmiller, “Learning an embedding space for transferable robot skills,” in
2018
Cited alongside, same era.
I. Popov, N. Heess, T. Lillicrap, R. Hafner, G. Barth-Maron, M. Vecerik, T. Lampe, Y. Tassa, T. Erez, and M. Riedmiller, “Data-efficient deep reinforcement learning for dexterous manipulation,” 2018
2018
Cited alongside, same era.
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine, “Learning complex dexterous manipulation with deep reinforcement learning and demonstrations,” in
2018
Cited alongside, same era.
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y. Bai, D. Hafner, S. Bohez, and V. Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” in
2018
Cited alongside, same era.
R. Martín-Martín, M. A. Lee, R. Gardner, S. Savarese, J. Bohg, and A. Garg, “Variable impedance control in end-effector space: An action space for reinforcement learning in contact-rich tasks,” in
2019
Later among the works it cites.
C. Nash and C. Durkan, “Autoregressive energy machines,” in
2019
Later among the works it cites.
N. Rahaman, A. Baratin, D. Arpit, F. Draxler, M. Lin, F. Hamprecht, Y. Bengio, and A. Courville, “On the spectral bias of neural networks,” in
2019
Later among the works it cites.
B. Ronen, D. Jacobs, Y. Kasten, and S. Kritchman, “The convergence rate of neural networks for learned functions of different frequencies,”
2019
Later among the works it cites.
T. Silver, K. Allen, J. Tenenbaum, and L. Kaelbling, “Residual policy learning,” 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
J. Viereck, J. Kozolinsky, A. Herzog, and L. Righetti, “Learning a structured neural network policy for a hopping task,” vol. 3, no. 4. Institute of Electrical and Electronics Engineers (IEEE), Oct 2018, p. 4092–4099. [Online]. Available: http://dx.doi.org/10.1109/LRA.2018.2861466
2018
Cited alongside, same era.
R. Basri, D. Jacobs, Y. Kasten, and S. Kritchman, “The convergence rate of neural networks for learned functions of different frequencies,” in
2019
Cited alongside, same era.
A. Bietti and J. Mairal, “On the inductive bias of neural tangent kernels,”
2019
Cited alongside, same era.
Y. Du and I. Mordatch, “Implicit generation and modeling with energy based models,” 2019
2019
Cited alongside, same era.
T. Johannink, S. Bahl, A. Nair, J. Luo, A. Kumar, M. Loskyll, J. A. Ojea, E. Solowjow, and S. Levine, “Residual reinforcement learning for robot control,” in
2019
Cited alongside, same era.
R. Martín-Martín, M. Lee, R. Gardner, S. Savarese, J. Bohg, and A. Garg, “Variable impedance control in end-effector space. an action space for reinforcement learning in contact rich tasks,” in
2019
Cited alongside, same era.
P. Varin, L. Grossman, and S. Kuindersma, “A comparison of action spaces for learning manipulation tasks,” in
2019
Later among the works it cites.
2020
Later among the works it cites.
M. Pflueger and G. Sukhatme, “Plan-space state embeddings for improved reinforcement learning,”
2020
Later among the works it cites.
T. Yu, S. Kumar, A. Gupta, S. Levine, K. Hausman, and C. Finn, “Gradient surgery for multi-task learning,” in
2020
Later among the works it cites.
H. Duan, J. Dao, K. Green, T. Apgar, A. Fern, and J. Hurst, “Learning task space actions for bipedal locomotion,” in
2021
Later among the works it cites.