Fetching the paper…
Reading the bibliography…
Learning to solve complex manipulation tasks from visual observations is a dominant challenge for real-world robot learning.
F. Kaplan, P.-Y. Oudeyer, E. Kubinyi, and A. Miklósi, “Robotic clicker training,”
2002
Earlier work this paper cites.
B. Blumberg, M. Downie, Y. Ivanov, M. Berlin, M. P. Johnson, and B. Tomlinson, “Integrated learning for interactive synthetic characters,” in
2002
Earlier work this paper cites.
W. B. Knox and P. Stone, “Tamer: Training an agent manually via evaluative reinforcement,” in
2008
Earlier work this paper cites.
——, “Interactively shaping agents via human reinforcement: The tamer framework,” in
2009
Earlier work this paper cites.
S. Chernova and M. Veloso, “Interactive policy learning through confidence-based autonomy,”
2009
Earlier work this paper cites.
P. Abbeel, A. Coates, and A. Y. Ng, “Autonomous helicopter aerobatics through apprenticeship learning,”
2010
Earlier work this paper cites.
C. Meriçli, M. Veloso, and H. L. Akin, “Complementary humanoid behavior shaping using corrective demonstration,” in
2010
Earlier work this paper cites.
M. Zucker, N. Ratliff, M. Stolle, J. Chestnutt, J. A. Bagnell, C. G. Atkeson, and J. Kuffner, “Optimization and learning for rough terrain legged locomotion,”
2011
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in
2011
Earlier work this paper cites.
R. Akrour, M. Schoenauer, and M. Sebag, “Preference-based policy learning,” in
2011
Earlier work this paper cites.
S. Griffith, K. Subramanian, J. Scholz, C. L. Isbell, and A. L. Thomaz, “Policy shaping: Integrating human feedback with reinforcement learning,” in
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
T. Osa, N. Sugita, and M. Mitsuishi, “Online trajectory planning in dynamic environments for surgical task automation.” in
2014
Earlier work this paper cites.
M. Schoenauer, R. Akrour, M. Sebag, and J.-C. Souplet, “Programming by feedback,” in
2014
Earlier work this paper cites.
T. Cederborg, I. Grover, C. L. Isbell Jr, and A. L. Thomaz, “Policy shaping with human teachers.” in
2015
Cited alongside, same era.
C. Celemin and J. Ruiz-del Solar, “Coach: Learning continuous actions from corrective advice communicated by humans,” in
2015
Cited alongside, same era.
A. Jain, S. Sharma, T. Joachims, and A. Saxena, “Learning preferences for manipulation tasks from online coactive feedback,”
2015
Cited alongside, same era.
J. MacGlashan, M. K. Ho, R. Loftin, B. Peng, G. Wang, D. L. Roberts, M. E. Taylor, and M. L. Littman, “Interactive learning from policy-dependent human feedback,” in
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2019
Later among the works it cites.
M. Kelly, C. Sidrane, K. Driggs-Campbell, and M. J. Kochenderfer, “Hg-dagger: Interactive imitation learning with human experts,” in
2019
Later among the works it cites.
R. Pérez-Dattari, C. Celemin, J. Ruiz-del Solar, and J. Kober, “Continuous control for high-dimensional state spaces: An interactive learning approach,” in
2019
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
G. Warnell, N. Waytowich, V. Lawhern, and P. Stone, “Deep tamer: Interactive agent shaping in high-dimensional state spaces,” in
2018
Cited alongside, same era.
Y. Cui and S. Niekum, “Active reward learning from critiques,” in
2018
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in
2018
Cited alongside, same era.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,”
2019
Cited alongside, same era.
2019
Later among the works it cites.
A. Zeng, S. Song, J. Lee, A. Rodriguez, and T. Funkhouser, “Tossingbot: Learning to throw arbitrary objects with residual physics,”
2020
Later among the works it cites.
W. Burgard, A. Valada, N. Radwan, T. Naseer, J. Zhang, J. Vertens, O. Mees, A. Eitel, and G. Oliveira, “Perspectives on deep multimodel robot learning,” in
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
S. James, Z. Ma, D. Rovick Arrojo, and A. J. Davison, “Rlbench: The robot learning benchmark & learning environment,”
2020
Later among the works it cites.
E. Chisari, A. Liniger, A. Rupenyan, L. Van Gool, and J. Lygeros, “Learning from simulation, racing in reality,” in
2021
Closest in time.
2021
Closest in time.
J. V. Hurtado, L. Londoño, and A. Valada, “From learning to relearning: A framework for diminishing bias in social robot navigation,”
2021
Closest in time.